Ad hoc autonomous agent teams: Collaboration without pre-coordination
Peter Stone, Gal Kaminka, Sarit Kraus, and Jeffrey Rosenschein. 2010 · 2010
Earlier work this paper cites.
Aateam: Achieving the ad hoc teamwork by employing the attention mechanism
Shuo Chen, Ewa Andrejczuk, Zhiguang Cao, and Jie Zhang. 2020 · 2020
Earlier work this paper cites.
A penny for your thoughts: The value of communication in ad hoc teamwork
Reuth Mirsky, William Macke, Andy Wang, Harel Yedidsion, and Peter Stone. 2020 · 2020
Earlier work this paper cites.
Deep reinforcement learning with stacked hierarchical attention for text-based games
Yunqiu Xu, Meng Fang, Ling Chen, Yali Du, Joey Tianyi Zhou, and Chengqi Zhang. 2020 · 2020
Earlier work this paper cites.
Expected value of communication for planning in ad hoc teamwork
William Macke, Reuth Mirsky, and Peter Stone. 2021 · 2021
Earlier work this paper cites.
Towards open ad hoc teamwork using graph-based policy learning
Muhammad A Rahman, Niklas Hopner, Filippos Christianos, and Stefano V Albrecht. 2021 · 2021
Earlier work this paper cites.
Generalization in text-based games via hierarchical reinforcement learning
Yunqiu Xu, Meng Fang, Ling Chen, Yali Du, and Chengqi Zhang. 2021 · 2021
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Earlier work this paper cites.
A survey of ad hoc teamwork: Definitions, methods, and open problems
Reuth Mirsky, Ignacio Carlucho, Arrasy Rahman, Elliot Fosong, William Macke, Mohan Sridharan, Peter Stone, and Stefano V Albrecht. 2022 · 2022
Earlier work this paper cites.
Stay moral and explore: Learn to behave morally in text-based games
Zijing Shi, Meng Fang, Yunqiu Xu, Ling Chen, and Yali Du. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Earlier work this paper cites.