Fetching the paper…
Reading the bibliography…
The desire and ability to seek new information strategically are fundamental to human learning but often overlooked in current language agent evaluation.
Computational interpretations of the gricean maxims in the generation of referring expressions
Robert Dale and Ehud Reiter. 1995 · 1995
Earlier work this paper cites.
Policy invariance under reward transformations: Theory and application to reward shaping
Andrew Y Ng, Daishi Harada, and Stuart J Russell. 1999 · 1999
Earlier work this paper cites.
Information-oriented online shopping behavior in electronic commerce environment
Chun-An Chen. 2009 · 2009
Earlier work this paper cites.
Toward an integrated framework for online consumer behavior and decision making process: A review
William K Darley, Charles Blankson, and Denise J Luethge. 2010 · 2010
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
Jürgen Schmidhuber. 2010 · 2010
Earlier work this paper cites.
Towards information-seeking agents
Philip Bachman, Alessandro Sordoni, and Adam Trischler. 2016 · 2016
Earlier work this paper cites.
Deep reinforcement learning with a natural language action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf. 2016 · 2016
Earlier work this paper cites.
Towards end-to-end learning for dialog state tracking and management using deep reinforcement learning
Tiancheng Zhao and Maxine Eskenazi. 2016 · 2016
Earlier work this paper cites.
Learning cooperative visual dialog agents with deep reinforcement learning
Abhishek Das, Satwik Kottur, José M. F. Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Guesswhat?! visual object discovery through multi-modal dialogue
Harm de Vries, Florian Strub, Sarath Chandar, Olivier Pietquin, Hugo Larochelle, and Aaron C. Courville. 2017 · 2017
Cited alongside, same era.
Towards end-to-end reinforcement learning of dialogue agents for information access
Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, and Li Deng. 2017 · 2017
Cited alongside, same era.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Cited alongside, same era.
Stay hungry, stay focused: Generating informative and specific questions in information-seeking conversations
Peng Qi, Yuhao Zhang, and Christopher D. Manning. 2020 · 2020
Cited alongside, same era.
Interactive machine comprehension with information seeking agents
Xingdi Yuan, Jie Fu, Marc-Alexandre Côté, Yi Tay, Chris Pal, and Adam Trischler. 2020 · 2020
Cited alongside, same era.
Decision-oriented dialogue for human-ai collaboration
Jessy Lin, Nicholas Tomlin, Jacob Andreas, and Jason Eisner. 2023 · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Later among the works it cites.
Know your audience: specializing grounded language models with listener subtraction
Aaditya K Singh, David Ding, Andrew Saxe, Felix Hill, and Andrew Lampinen. 2023 · 2023
Later among the works it cites.
Task ambiguity in humans and language models
Alex Tamkin, Kunal Handa, Avash Shrestha, and Noah D. Goodman. 2023 · 2023
Later among the works it cites.
Mint: Evaluating llms in multi-turn interaction with tools and language feedback
Xingyao Wang, Zihan Wang, Jiateng Liu, Yangyi Chen, Lifan Yuan, Hao Peng, and Heng Ji. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Open-domain clarification question generation without question examples
Julia White, Gabriel Poesia, Robert Hawkins, Dorsa Sadigh, and Noah Goodman. 2021 · 2021
Cited alongside, same era.
Multi-stage episodic control for strategic exploration in text games
Jens Tuyls, Shunyu Yao, Sham Kakade, and Karthik Narasimhan. 2022 · 2022
Cited alongside, same era.
Webshop: Towards scalable real-world web interaction with grounded language agents
Shunyu Yao, Howard Chen, John Yang, and Karthik R Narasimhan. 2022 · 2022
Cited alongside, same era.
Conversational information seeking
Hamed Zamani, Johanne R Trippas, Jeff Dalton, and Filip Radlinski. 2022 · 2022
Cited alongside, same era.
Eliciting human preferences with language models
Belinda Z. Li, Alex Tamkin, Noah Goodman, and Jacob Andreas. 2023a
Cited in the paper.
Yuan Li, Yixuan Zhang, and Lichao Sun. 2023b
Cited in the paper.
Lost in the middle: How language models use long contexts
Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. 2023a
Cited in the paper.
Openagents: An open platform for language agents in the wild
Tianbao Xie, Fan Zhou, Zhoujun Cheng, Peng Shi, Luoxuan Weng, Yitao Liu, Toh Jing Hua, Junning Zhao, Qian Liu, Che Liu, Leo Z. Liu, Yiheng Xu, Hongjin Su, Dongchan Shin, Caiming Xiong, and Tao Yu. 2023 · 2023
Later among the works it cites.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik R. Narasimhan, and Yuan Cao. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric. P Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica. 2023 · 2023
Later among the works it cites.
Webarena: A realistic web environment for building autonomous agents
Shuyan Zhou, Frank F. Xu, Hao Zhu, Xuhui Zhou, Robert Lo, Abishek Sridhar, Xianyi Cheng, Tianyue Ou, Yonatan Bisk, Daniel Fried, Uri Alon, and Graham Neubig. 2023 · 2023
Later among the works it cites.