Fetching the paper…
Reading the bibliography…
Recent advancements in Large Language Models (LLMs) have led to significant breakthroughs in various natural language processing tasks.
Constructing a multi-hop qa dataset for comprehensive evaluation of reasoning steps
Ho, X.; Nguyen, A.-K. D.; Sugawara, S.; and Aizawa, A. 2020 · 2011
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Rajpurkar, P.; Zhang, J.; Lopyrev, K.; and Liang, P. 2016 · 2016
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Clark, P.; Cowhey, I.; Etzioni, O.; Khot, T.; Sabharwal, A.; Schoenick, C.; and Tafjord, O. 2018 · 2018
Earlier work this paper cites.
Wizard of Wikipedia: Knowledge-Powered Conversational Agents
Dinan, E.; Roller, S.; Shuster, K.; Fan, A.; Auli, M.; and Weston, J. 2018 · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Mihaylov, T.; Clark, P.; Khot, T.; and Sabharwal, A. 2018 · 2018
Earlier work this paper cites.
FEVER: a Large-scale Dataset for Fact Extraction and VERification
Thorne, J.; Vlachos, A.; Christodoulopoulos, C.; and Mittal, A. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Yang, Z.; Qi, P.; Zhang, S.; Bengio, Y.; Cohen, W.; Salakhutdinov, R.; and Manning, C. D. 2018 · 2018
Earlier work this paper cites.
Cosmos QA: Machine Reading Comprehension with Contextual Commonsense Reasoning
Huang, L.; Le Bras, R.; Bhagavatula, C.; and Choi, Y. 2019 · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Kwiatkowski, T.; Palomaki, J.; Redfield, O.; Collins, M.; Parikh, A.; Alberti, C.; Epstein, D.; Polosukhin, I.; Devlin, J.; Lee, K.; et al. 2019 · 2019
Earlier work this paper cites.
Dense Passage Retrieval for Open-Domain Question Answering
Karpukhin, V.; Oguz, B.; Min, S.; Lewis, P.; Wu, L.; Edunov, S.; Chen, D.; and Yih, W.-t. 2020 · 2020
Earlier work this paper cites.
Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters
Rasley, J.; Rajbhandari, S.; Ruwase, O.; and He, Y. 2020 · 2020
Earlier work this paper cites.
CoLAKE: Contextualized Language and Knowledge Embedding
Sun, T.; Shao, Y.; Qiu, X.; Guo, Q.; Hu, Y.; Huang, X.-J.; and Zhang, Z. 2020 · 2020
Earlier work this paper cites.
Open-Domain Question Answering Goes Conversational via Question Rewriting
Anantha, R.; Vakulenko, S.; Tu, Z.; Longpre, S.; Pulman, S.; and Chappidi, S. 2021 · 2021
Earlier work this paper cites.
Editing Factual Knowledge in Language Models
De Cao, N.; Aziz, W.; and Titov, I. 2021 · 2021
Earlier work this paper cites.
Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies
Geva, M.; Khashabi, D.; Segal, E.; Khot, T.; Roth, D.; and Berant, J. 2021 · 2021
Earlier work this paper cites.
LoRA: Low-Rank Adaptation of Large Language Models
Hu, E. J.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; Chen, W.; et al. 2021 · 2021
Earlier work this paper cites.
Unsupervised dense information retrieval with contrastive learning
Izacard, G.; Caron, M.; Hosseini, L.; Riedel, S.; Bojanowski, P.; Joulin, A.; and Grave, E. 2021 · 2021
Earlier work this paper cites.
PubHealthTab: A public health table-based dataset for evidence-based fact checking
Akhtar, M.; Cocarascu, O.; and Simperl, E. 2022 · 2022
Earlier work this paper cites.
Mallen, A.; Asai, A.; Zhong, V.; Das, R.; Khashabi, D.; and Hajishirzi, H. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Ouyang, L.; Wu, J.; Jiang, X.; Almeida, D.; Wainwright, C.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al. 2022 · 2022
Cited alongside, same era.
Large language models encode clinical knowledge
Singhal, K.; Azizi, S.; Tu, T.; Mahdavi, S. S.; Wei, J.; Chung, H. W.; Scales, N.; Tanwani, A.; Cole-Lewis, H.; Pfohl, S.; et al. 2022 · 2022
Cited alongside, same era.
ASQA: Factoid Questions Meet Long-Form Answers
Stelmakh, I.; Luan, Y.; Dhingra, B.; and Chang, M.-W. 2022 · 2022
Cited alongside, same era.
Black-box tuning for language-model-as-a-service
Sun, T.; Shao, Y.; Qian, H.; Huang, X.; and Qiu, X. 2022 · 2022
Cited alongside, same era.
Ra-dit: Retrieval-augmented dual instruction tuning
Lin, X. V.; Chen, X.; Chen, M.; Shi, W.; Lomeli, M.; James, R.; Rodriguez, P.; Kahn, J.; Szilvasy, G.; Lewis, M.; et al. 2023 · 2023
Later among the works it cites.
Llm+ p: Empowering large language models with optimal planning proficiency
Liu, B.; Jiang, Y.; Zhang, X.; Liu, Q.; Zhang, S.; Biswas, J.; and Stone, P. 2023 · 2023
Later among the works it cites.
The flan collection: Designing data and methods for effective instruction tuning
Longpre, S.; Hou, L.; Vu, T.; Webson, A.; Chung, H. W.; Tay, Y.; Zhou, D.; Le, Q. V.; Zoph, B.; Wei, J.; et al. 2023 · 2023
Later among the works it cites.
Query Rewriting in Retrieval-Augmented Large Language Models
Ma, X.; Gong, Y.; He, P.; Zhao, H.; and Duan, N. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Logic-Driven Context Extension and Data Augmentation for Logical Reasoning of Text
Wang, S.; Zhong, W.; Tang, D.; Wei, Z.; Fan, Z.; Jiang, D.; Zhou, M.; and Duan, N. 2022b · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Xia, F.; Chi, E.; Le, Q. V.; Zhou, D.; et al. 2022 · 2022
Cited alongside, same era.
Achiam, J.; Adler, S.; Agarwal, S.; Ahmad, L.; Akkaya, I.; Aleman, F. L.; Almeida, D.; Altenschmidt, J.; Altman, S.; Anadkat, S.; et al. 2023 · 2023
Cited alongside, same era.
Self-rag: Learning to retrieve, generate, and critique through self-reflection
Asai, A.; Wu, Z.; Wang, Y.; Sil, A.; and Hajishirzi, H. 2023 · 2023
Cited alongside, same era.
Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks
Chen, W.; Ma, X.; Wang, X.; and Cohen, W. W. 2023 · 2023
Cited alongside, same era.
Enabling Large Language Models to Generate Text with Citations
Gao, T.; Yen, H.; Yu, J.; and Chen, D. 2023 · 2023
Cited alongside, same era.
Metagpt: Meta programming for multi-agent collaborative framework
Hong, S.; Zheng, X.; Chen, J.; Cheng, Y.; Wang, J.; Zhang, C.; Wang, Z.; Yau, S. K. S.; Lin, Z.; Zhou, L.; et al. 2023 · 2023
Cited alongside, same era.
Peng, B.; Li, C.; He, P.; Galley, M.; and Gao, J. 2023 · 2023
Later among the works it cites.
Replug: Retrieval-augmented black-box language models
Shi, W.; Min, S.; Yasunaga, M.; Seo, M.; James, R.; Lewis, M.; Zettlemoyer, L.; and Yih, W.-t. 2023 · 2023
Later among the works it cites.
Llm-planner: Few-shot grounded planning for embodied agents with large language models
Song, C. H.; Wu, J.; Washington, C.; Sadler, B. M.; Chao, W.-L.; and Su, Y. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron, H.; Martin, L.; Stone, K.; Albert, P.; Almahairi, A.; Babaei, Y.; Bashlykov, N.; Batra, S.; Bhargava, P.; Bhosale, S.; et al. 2023 · 2023
Later among the works it cites.
Self-supervised cross-view representation reconstruction for change captioning
Tu, Y.; Li, L.; Su, L.; Zha, Z.-J.; Yan, C.; and Huang, Q. 2023 · 2023
Later among the works it cites.
Learning to filter context for retrieval-augmented generation
Wang, Z.; Araki, J.; Jiang, Z.; Parvez, M. R.; and Neubig, G. 2023 · 2023
Later among the works it cites.
Recomp: Improving retrieval-augmented lms with compression and selective augmentation
Xu, F.; Shi, W.; and Choi, E. 2023 · 2023
Later among the works it cites.
ReAct: Synergizing Reasoning and Acting in Language Models
Yao, S.; Zhao, J.; Yu, D.; Du, N.; Shafran, I.; Narasimhan, K.; and Cao, Y. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Zheng, L.; Chiang, W.-L.; Sheng, Y.; Zhuang, S.; Wu, Z.; Zhuang, Y.; Lin, Z.; Li, Z.; Li, D.; Xing, E.; et al. 2023 · 2023
Later among the works it cites.
Fine-tuned large language model for visualization system: A study on self-regulated learning in education
Gao, L.; Lu, J.; Shao, Z.; Lin, Z.; Yue, S.; Ieong, C.; Sun, Y.; Zauner, R. J.; Wei, Z.; and Chen, S. 2024 · 2024
Closest in time.
Openassistant conversations-democratizing large language model alignment
Köpf, A.; Kilcher, Y.; von Rütte, D.; Anagnostidis, S.; Tam, Z. R.; Stevens, K.; Barhoum, A.; Nguyen, D.; Stanley, O.; Nagyfi, R.; et al. 2024 · 2024
Closest in time.
Mou, X.; Wei, Z.; and Huang, X. 2024 · 2024
Closest in time.
Small llms are weak tool learners: A multi-llm agent
Shen, W.; Li, C.; Chen, H.; Yan, M.; Quan, X.; Chen, H.; Zhang, J.; and Huang, F. 2024 · 2024
Closest in time.
LawLLM: Intelligent Legal System with Legal Reasoning and Verifiable Retrieval
Yue, S.; Liu, S.; Zhou, Y.; Shen, C.; Wang, S.; Xiao, Y.; Li, B.; Song, Y.; Shen, X.; Chen, W.; et al. 2024 · 2024
Closest in time.