Fetching the paper…
Reading the bibliography…
Making the content generated by Large Language Model (LLM), accurate, credible and traceable is crucial, especially in complex knowledge-intensive tasks that require multi-step reasoning and each step needs knowledge to solve.
Depth-first search and linear graph algorithms. In 12th Annual Symposium on Switching and Automata Theory (swat 1971) . 114–121
Robert Tarjan. 1971 · 1971
Earlier work this paper cites.
REALM: Retrieval-Augmented Language Model Pre-Training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang. 2020 · 2002
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries. In Text Summarization Branches Out . Association for Computational Linguistics, Barcelona, Spain, 74–81
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Leveraging Passage Retrieval with Generative Models for Open Domain Question Answering
Gautier Izacard and Edouard Grave. 2020 · 2007
Earlier work this paper cites.
Knowledge-Aware Language Model Pretraining
Corby Rosset, Chenyan Xiong, Minh Phan, et al · 2007
Earlier work this paper cites.
Uncovering the community structure associated with the diffusion dynamics on networks
Xue-Qi Cheng and Hua-Wei Shen. 2010 · 2010
Earlier work this paper cites.
Top-k learning to rank: labeling, ranking and evaluation. In Proceedings of the 35th international ACM SIGIR conference on Research and development in information retrieval . 751–760
Shuzi Niu, Jiafeng Guo, Yanyan Lan, and Xueqi Cheng. 2012 · 2012
Earlier work this paper cites.
Zero-Shot Relation Extraction via Reading Comprehension. In Proceedings of the 2017 Conference on CoNLL) . Association for Computational Linguistics, Vancouver, Canada, 333–342
Omer Levy, Minjoon Seo, Eunsol Choi, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
T-REx: A Large Scale Alignment of Natural Language with Knowledge Base Triples. In Proceedings of the 2018 Conference on LREC . European Language Resources Association (ELRA), Miyazaki, Japan
Hady Elsahar, Pavlos Vougiouklis, Arslen Remaci, et al · 2018
Earlier work this paper cites.
FEVER: a Large-scale Dataset for Fact Extraction and VERification. In Proceedings of the 2018 Conference on NAACL . Association for Computational Linguistics, New Orleans, Louisiana, 809–819
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering. In Proceedings of the 2018 Conference on EMNLP . Association for Computational Linguistics, Brussels, Belgium, 2369–2380
Zhilin Yang, Peng Qi, Saizheng Zhang, et al · 2018
Earlier work this paper cites.
ELI5: Long Form Question Answering. In Proceedings of the 2019 Conference on ACL . Association for Computational Linguistics, Florence, Italy, 3558–3567
Angela Fan, Yacine Jernite, Ethan Perez, et al · 2019
Earlier work this paper cites.
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps. In Proceedings of the 2020 Conference on COLING . International Committee on Computational Linguistics, Barcelona, Spain (Online), 6609–6625
Xanh Ho, Anh-Khoa Duong Nguyen, Saku Sugawara, and Akiko Aizawa. 2020 · 2020
Earlier work this paper cites.
Dense Passage Retrieval for Open-Domain Question Answering. In Proceedings of the 2020 Conference on EMNLP . Association for Computational Linguistics, Online, 6769–6781
Vladimir Karpukhin, Barlas Oguz, Sewon Min, et al · 2020
Earlier work this paper cites.
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks. In Proceedings of the 2020 Conference on NeurIPS
Patrick S. H. Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2020 · 2020
Earlier work this paper cites.
Did Aristotle Use a Laptop? A Question Answering Benchmark with Implicit Reasoning Strategies
Mor Geva, Daniel Khashabi, Elad Segal, Tushar Khot, Dan Roth, and Jonathan Berant. 2021 · 2021
Cited alongside, same era.
Narrative Question Answering with Cutting-Edge Open-Domain QA Techniques: A Comprehensive Study
Xiangyang Mou, Chenghao Yang, Mo Yu, Bingsheng Yao, Xiaoxiao Guo, Saloni Potdar, and Hui Su. 2021 · 2021
Cited alongside, same era.
KILT: a Benchmark for Knowledge Intensive Language Tasks. In Proceedings of the 2021 Conference on NAACL . Association for Computational Linguistics, Online, 2523–2544
Fabio Petroni, Aleksandra Piktus, Angela Fan, et al · 2021
Cited alongside, same era.
Adaptive information seeking for open-domain question answering
Yunchang Zhu, Liang Pang, Yanyan Lan, Huawei Shen, and Xueqi Cheng. 2021 · 2021
Cited alongside, same era.
Least-to-Most Prompting Enables Complex Reasoning in Large Language Models
Denny Zhou, Nathanael Schärli, Le Hou, et al · 2022
Later among the works it cites.
Large language models and the perils of their hallucinations
Razvan Azamfirei, Sapna R Kudchadkar, and James Fackler. 2023 · 2023
Closest in time.
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, and Pascale Fung. 2023 · 2023
Closest in time.
Compositional Semantic Parsing with Large Language Models. In The Eleventh International Conference on Learning Representations
Andrew Drozdov, Nathanael Schärli, Ekin Akyürek, Nathan Scales, Xinying Song, Xinyun Chen, Olivier Bousquet, and Denny Zhou. 2023 · 2023
Closest in time.
Complexity-Based Prompting for Multi-step Reasoning. In The Eleventh International Conference on Learning Representations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, et al · 2022
Cited alongside, same era.
Large Language Models Struggle to Learn Long-Tail Knowledge
Nikhil Kandpal, Haikang Deng, Adam Roberts, Eric Wallace, and Colin Raffel. 2022 · 2022
Cited alongside, same era.
Large Language Models are Zero-Shot Reasoners. In Advances in Neural Information Processing Systems , Alice H. Oh, Alekh Agarwal, Danielle Belgrave, and Kyunghyun Cho (Eds.)
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Cited alongside, same era.
Nonparametric Masked Language Modeling
Sewon Min, Weijia Shi, Mike Lewis, Xilun Chen, Wen-tau Yih, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2022 · 2022
Cited alongside, same era.
ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction. In Proceedings of the 2022 Conference on NAACL . Association for Computational Linguistics, Seattle, United States, 3715–3734
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon, Christopher Potts, and Matei Zaharia. 2022 · 2022
Cited alongside, same era.
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R. Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, et al · 2022
Cited alongside, same era.
MuSiQue: Multihop Questions via Single-hop Question Composition
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022 · 2022
Cited alongside, same era.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed H. Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Cited alongside, same era.
Yao Fu, Hao Peng, Ashish Sabharwal, Peter Clark, and Tushar Khot. 2023 · 2023
Closest in time.
Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP
Omar Khattab, Keshav Santhanam, Xiang Lisa Li, et al · 2023
Closest in time.
Measuring and Narrowing the Compositionality Gap in Language Models
Ofir Press, Muru Zhang, Sewon Min, Ludwig Schmidt, Noah A. Smith, and Mike Lewis. 2023 · 2023
Closest in time.
Hongjing Qian, Yutao Zhu, Zhicheng Dou, et al · 2023
Closest in time.
Toolformer: Language Models Can Teach Themselves to Use Tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Closest in time.
Recitation-Augmented Language Models. In The Eleventh International Conference on Learning Representations
Zhiqing Sun, Xuezhi Wang, Yi Tay, Yiming Yang, and Denny Zhou. 2023 · 2023
Closest in time.
Shicheng Xu, Liang Pang, Huawei Shen, and Xueqi Cheng. 2023 · 2023
Closest in time.
Automatic Chain of Thought Prompting in Large Language Models. In The Eleventh International Conference on Learning Representations
Zhuosheng Zhang, Aston Zhang, Mu Li, and Alex Smola. 2023 · 2023
Closest in time.
Verify-and-Edit: A Knowledge-Enhanced Chain-of-Thought Framework. In Proceedings of the 2023 Conference on ACL . Association for Computational Linguistics, Toronto, Canada, 5823–5840
Ruochen Zhao, Xingxuan Li, Shafiq Joty, Chengwei Qin, and Lidong Bing. 2023a · 2023
Closest in time.
List-aware Reranking-Truncation Joint Model for Search and Retrieval-augmented Generation
Shicheng Xu, Liang Pang, Jun Xu, Huawei Shen, and Xueqi Cheng. 2024 · 2024
Closest in time.