Fetching the paper…
Reading the bibliography…
Multi-agent strategies have emerged as a promising approach to enhance the reasoning abilities of Large Language Models (LLMs) by assigning specialized roles in the problem-solving process.
Rankgen: Improving text generation with large ranking models, 2022
Kalpesh Krishna, Yapei Chang, John Wieting, and Mohit Iyyer · 2022
Earlier work this paper cites.
Faithfulness-aware decoding strategies for abstractive summarization, 2023
David Wan, Mengwen Liu, Kathleen McKeown, Markus Dreyer, and Mohit Bansal · 2023
Earlier work this paper cites.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan · 2023
Earlier work this paper cites.
Parsel: Algorithmic reasoning with language models by composing decompositions
Eric Zelikman, Qian Huang, Gabriel Poesia, Noah Goodman, and Nick Haber · 2023
Cited alongside, same era.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Amy Yang, Angela Fan, et al · 2024
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman
Cited in the paper.
Improving factuality and reasoning in language models through multiagent debate
Yilun Du, Shuang Li, Antonio Torralba, Joshua B. Tenenbaum, and Igor Mordatch
Cited in the paper.
Counterfactual debating with preset stances for hallucination elimination of LLMs
Yi Fang, Moxin Li, Wenjie Wang, Hui Lin, and Fuli Feng
Cited in the paper.
HaluEval: A large-scale hallucination evaluation benchmark for large language models
Junyi Li, Xiaoxue Cheng, Wayne Xin Zhao, Jian-Yun Nie, and Ji-Rong Wen
Cited in the paper.
Nousresearch/meta-llama-3.1-70b
NousResearch
Cited in the paper.
Nousresearch/meta-llama-3.1-8b
NousResearch
Cited in the paper.
GPT-3.5 Turbo
OpenAI
Cited in the paper.
Multi-agent collaboration: Harnessing the power of intelligent LLM agents
Yashar Talebirad and Amirhossein Nadiri
Cited in the paper.
Ziyi Tang, Ruilin Wang, Weixing Chen, Keze Wang, Yang Liu, Tianshui Chen, and Liang Lin
Cited in the paper.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou
Cited in the paper.
Large language model based multi-agents: A survey of progress and challenges
Taicheng Guo, Xiuying Chen, Yaqi Wang, Ruidi Chang, Shichao Pei, Nitesh V. Chawla, Olaf Wiest, and Xiangliang Zhang · 2024
Closest in time.
GPT-4o-mini
OpenAI · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…