Fetching the paper…
Reading the bibliography…
With the rise of Large Language Models(LLMs), it has become crucial to understand their capabilities and limitations in deciphering and explaining the complex web of causal relationships that language entails.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Lan, Z.; Chen, M.; Goodman, S.; Gimpel, K.; Sharma, P.; and Soricut, R. 2019 · 1909
Earlier work this paper cites.
Counterfactual story reasoning and generation
Qin, L.; Bosselut, A.; Holtzman, A.; Bhagavatula, C.; Clark, E.; and Choi, Y. 2019 · 1909
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Conneau, A.; Khandelwal, K.; Goyal, N.; Chaudhary, V.; Wenzek, G.; Guzmán, F.; Grave, E.; Ott, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1911
Earlier work this paper cites.
Deberta: Decoding-enhanced bert with disentangled attention
He, P.; Liu, X.; Gao, J.; and Chen, W. 2020 · 2006
Earlier work this paper cites.
ConceptNet 5.5: An Open Multilingual Graph of General Knowledge
Speer, R.; Chin, J.; and Havasi, C. 2017 · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2018 · 2018
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2020 · 2020
Earlier work this paper cites.
Causal direction of data collection matters: Implications of causal and anticausal learning for NLP
Jin, Z.; von Kügelgen, J.; Ni, J.; Vaidhya, T.; Kaushal, A.; Sachan, M.; and Schoelkopf, B. 2021 · 2021
Cited alongside, same era.
COM2SENSE: A Commonsense Reasoning Benchmark with Complementary Sentences
Singh, S.; Wen, N.; Hou, Y.; Alipoormolabashi, P.; Wu, T.-L.; Ma, X.; and Peng, N. 2021 · 2021
Cited alongside, same era.
e-CARE: a New Dataset for Exploring Explainable Causal Reasoning
Du, L.; Ding, X.; Xiong, K.; Liu, T.; and Qin, B. 2022 · 2022
Cited alongside, same era.
Jiang, A. Q.; Sablayrolles, A.; Mensch, A.; Bamford, C.; Chaplot, D. S.; Casas, D. d. l.; Bressand, F.; Lengyel, G.; Lample, G.; Saulnier, L.; et al. 2023 · 2023
Cited alongside, same era.
Gemini: a family of highly capable multimodal models
Team, G.; Anil, R.; Borgeaud, S.; Wu, Y.; Alayrac, J.-B.; Yu, J.; Soricut, R.; Schalkwyk, J.; Dai, A. M.; Hauth, A.; et al. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H.; Martin, L.; Stone, K.; Albert, P.; Almahairi, A.; Babaei, Y.; Bashlykov, N.; Batra, S.; Bhargava, P.; Bhosale, S.; et al. 2023 · 2023
Later among the works it cites.
Large language models are better reasoners with self-verification
Weng, Y.; Zhu, M.; Xia, F.; Li, B.; He, S.; Liu, S.; Sun, B.; Liu, K.; and Zhao, J. 2023 · 2023
Later among the works it cites.
Causal parrots: Large language models may talk causality but are not causal
Zečević, M.; Willig, M.; Dhami, D. S.; and Kersting, K. 2023 · 2023
Later among the works it cites.
Understanding Causality with Large Language Models: Feasibility and Opportunities
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jin, Z.; Chen, Y.; Leeb, F.; Gresele, L.; Kamal, O.; Lyu, Z.; Blin, K.; Gonzalez Adauto, F.; Kleiman-Weiner, M.; Sachan, M.; and Schölkopf, B. 2023a · 2023
Cited alongside, same era.
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality
Kıçıman, E.; Ness, R.; Sharma, A.; and Tan, C. 2023 · 2023
Cited alongside, same era.
Leveraging Large Language Models for Topic Classification in the Domain of Public Affairs
Peña, A.; Morales, A.; Fierrez, J.; Serna, I.; Ortega-Garcia, J.; Puente, I.; Cordova, J.; and Cordova, G. 2023 · 2023
Cited alongside, same era.
Crab: Assessing the strength of causal relationships between real-world events
Romanou, A.; Montariol, S.; Paul, D.; Laugier, L.; Aberer, K.; and Bosselut, A. 2023 · 2023
Cited alongside, same era.
Can large language models infer causation from correlation?
Jin, Z.; Liu, J.; Lyu, Z.; Poff, S.; Sachan, M.; Mihalcea, R.; Diab, M.; and Schölkopf, B. 2023b
Cited in the paper.
Zhang, C.; Bauer, S.; Bennett, P.; Gao, J.; Gong, W.; Hilmkil, A.; Jennings, J.; Ma, C.; Minka, T.; Pawlowski, N.; and Vaughan, J. 2023 · 2023
Later among the works it cites.
Through the Lens of Core Competency: Survey on Evaluation of Large Language Models
Zhuang, Z.; Chen, Q.; Ma, L.; Li, M.; Han, Y.; Qian, Y.; Bai, H.; Feng, Z.; Zhang, W.; and Liu, T. 2023 · 2023
Later among the works it cites.
https://platform.openai.com/docs
OpenAI. 2024 · 2024
Closest in time.