Fetching the paper…
Reading the bibliography…
This paper considers the challenges Large Language Models (LLMs) face when reasoning over text that includes information involving uncertainty explicitly quantified via probability values.
Language models are few-shot learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Prolog Programming for Artificial Intelligence
Bratko, I. 2000 · 2000
Earlier work this paper cites.
ProbLog: A Probabilistic Prolog and Its Application in Link Discovery
De Raedt, L.; Kimmig, A.; and Toivonen, H. 2007 · 2007
Earlier work this paper cites.
Bayesian Rationality: The probabilistic approach to human reasoning
Oaksford, M.; and Chater, N. 2007 · 2007
Earlier work this paper cites.
Probabilistic Graphical Models: Principles and Techniques
Koller, D.; and Friedman, N. 2009 · 2009
Earlier work this paper cites.
Précis of Bayesian Rationality: The Probabilistic Approach to Human Reasoning
Oaksford, M.; and Chater, N. 2009 · 2009
Earlier work this paper cites.
Action formation and its epistemic (and other) backgrounds
Heritage, J. 2013 · 2013
Earlier work this paper cites.
pgmpy: Probabilistic graphical models using python
Ankan, A.; and Panda, A. 2015 · 2015
Earlier work this paper cites.
Whose decision? Negotiating epistemic and deontic rights in medical treatment decisions
Landmark, A. M. D.; Gulbrandsen, P.; and Svennevig, J. 2015 · 2015
Earlier work this paper cites.
Solving Probability Problems in Natural Language
Dries, A.; Kimmig, A.; Davis, J.; Belle, V.; and de Raedt, L. 2017 · 2017
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Earlier work this paper cites.
Uncertain words, uncertain texts. perception and effects of uncertainty in biomedical communication
Poggi, I.; D’Errico, F.; Vincze, L.; et al. 2019 · 2019
Earlier work this paper cites.
Training verifiers to solve math word problems
Cobbe, K.; Kosaraju, V.; Bavarian, M.; Chen, M.; Jun, H.; Kaiser, L.; Plappert, M.; Tworek, J.; Hilton, J.; Nakano, R.; et al. 2021 · 2021
Cited alongside, same era.
DomiKnowS: A Library for Integration of Symbolic Domain Knowledge in Deep Learning
Rajaby Faghihi, H.; Guo, Q.; Uszok, A.; Nafar, A.; and Kordjamshidi, P. 2021 · 2021
Cited alongside, same era.
RuleBERT: Teaching Soft Rules to Pre-Trained Language Models
Saeed, M.; Ahmadi, N.; Nakov, P.; and Papotti, P. 2021 · 2021
Cited alongside, same era.
Mapping probability word problems to executable representations
Suster, S.; Fivez, P.; Totis, P.; Kimmig, A.; Davis, J.; de Raedt, L.; and Daelemans, W. 2021 · 2021
Cited alongside, same era.
NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks
Mishra, S.; Mitra, A.; Varshney, N.; Sachdeva, B.; Clark, P.; Baral, C.; and Kalyan, A. 2022 · 2022
Cited alongside, same era.
Jiang, A. Q.; Sablayrolles, A.; Mensch, A.; Bamford, C.; Chaplot, D. S.; de las Casas, D.; Bressand, F.; Lengyel, G.; Lample, G.; Saulnier, L.; Lavaud, L. R.; Lachaux, M.-A.; Stock, P.; Scao, T. L.; Lavril, T.; Wang, T.; Lacroix, T.; and Sayed, W. E. 2023 · 2023
Later among the works it cites.
CLadder: A Benchmark to Assess Causal Reasoning Capabilities of Language Models
Jin, Z.; Chen, Y.; Leeb, F.; Gresele, L.; Kamal, O.; LYU, Z.; Blin, K.; Adauto, F. G.; Kleiman-Weiner, M.; Sachan, M.; and Schölkopf, B. 2023 · 2023
Later among the works it cites.
It Ain’t Over: A Multi-aspect Diverse Math Word Problem Dataset
Kim, J.; Kim, Y.; Baek, I.; Bak, J.; and Lee, J. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
ThinkSum: Probabilistic reasoning over sets using large language models
Ozturkler, B.; Malkin, N.; Wang, Z.; and Jojic, N. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Training language models to follow instructions with human feedback
Ouyang, L.; Wu, J.; Jiang, X.; Almeida, D.; Wainwright, C.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al. 2022 · 2022
Cited alongside, same era.
Stepgame: A new benchmark for robust multi-hop spatial reasoning in texts
Shi, Z.; Zhang, Q.; and Lipani, A. 2022 · 2022
Cited alongside, same era.
Sparks of artificial general intelligence: Early experiments with gpt-4
Bubeck, S.; Chandrasekaran, V.; Eldan, R.; Gehrke, J.; Horvitz, E.; Kamar, E.; Lee, P.; Lee, Y. T.; Li, Y.; Lundberg, S.; et al. 2023 · 2023
Cited alongside, same era.
Mathematical Capabilities of ChatGPT
Frieder, S.; Pinchetti, L.; Chevalier, A.; Griffiths, R.-R.; Salvatori, T.; Lukasiewicz, T.; Petersen, P. C.; and Berner, J. 2023 · 2023
Cited alongside, same era.
PAL: program-aided language models
Gao, L.; Madaan, A.; Zhou, S.; Alon, U.; Liu, P.; Yang, Y.; Callan, J.; and Neubig, G. 2023 · 2023
Cited alongside, same era.
Google Gemini AI
Google. 2023 · 2023
Cited alongside, same era.
Solving math word problems by combining language models with symbolic solvers
He-Yueya, J.; Poesia, G.; Wang, R. E.; and Goodman, N. D. 2023 · 2023
Cited alongside, same era.
Later among the works it cites.
Certified Deductive Reasoning with Language Models
Poesia, G.; Gandhi, K.; Zelikman, E.; and Goodman, N. D. 2023 · 2023
Later among the works it cites.
GLUECons: A Generic Benchmark for Learning under Constraints
Rajaby Faghihi, H.; Nafar, A.; Zheng, C.; Mirzaee, R.; Zhang, Y.; Uszok, A.; Wan, A.; Premsri, T.; Roth, D.; and Kordjamshidi, P. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H.; Martin, L.; Stone, K.; Albert, P.; Almahairi, A.; Babaei, Y.; Bashlykov, N.; Batra, S.; Bhargava, P.; Bhosale, S.; et al. 2023 · 2023
Later among the works it cites.
Llama 3 Model Card
AI@Meta. 2024 · 2024
Closest in time.
Teaching Probabilistic Logical Reasoning to Transformers
Nafar, A.; Venable, K. B.; and Kordjamshidi, P. 2024 · 2024
Closest in time.
Code Llama: Open Foundation Models for Code
Rozière, B.; Gehring, J.; Gloeckle, F.; Sootla, S.; Gat, I.; Tan, X. E.; Adi, Y.; Liu, J.; Sauvestre, R.; Remez, T.; Rapin, J.; Kozhevnikov, A.; Evtimov, I.; Bitton, J.; Bhatt, M.; Ferrer, C. C.; Grattafiori, A.; Xiong, W.; Défossez, A.; Copet, J.; Azhar, F.; Touvron, H.; Martin, L.; Usunier, N.; Scialom, T.; and Synnaeve, G. 2024 · 2024
Closest in time.
Toolformer: Language models can teach themselves to use tools
Schick, T.; Dwivedi-Yu, J.; Dessì, R.; Raileanu, R.; Lomeli, M.; Hambro, E.; Zettlemoyer, L.; Cancedda, N.; and Scialom, T. 2024 · 2024
Closest in time.