FANToM: A benchmark for stress-testing machine theory of mind in interactions
Hyunwoo Kim, Melanie Sclar, Xuhui Zhou, Ronan Bras, Gunhee Kim, Yejin Choi, and Maarten Sap. 2023b · 2023
Later among the works it cites.
Encouraging divergent thinking in large language models through multi-agent debate
Original
Tian Liang, Zhiwei He, Wenxiang Jiao, Xing Wang, Yan Wang, Rui Wang, Yujiu Yang, Zhaopeng Tu, and Shuming Shi. 2023 · 2023
Later among the works it cites.
Evaluating statistical language models as pragmatic reasoners
Benjamin Lipkin, Lionel Wong, Gabriel Grand, and Joshua B Tenenbaum. 2023 · 2023
Later among the works it cites.
Reproducibility in nlp: What have we learned from the checklist?
Ian H. Magnusson, Noah A. Smith, and Jesse Dodge. 2023 · 2023
Later among the works it cites.
Designing and detecting lies by reasoning about other agents
Lauren A Oey, Adena Schachner, and Edward Vul. 2023 · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph C. O’Brien, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein. 2023 · 2023
Later among the works it cites.
The goldilocks of pragmatic understanding: Fine-tuning strategy matters for implicature resolution by LLMs
Laura Eline Ruis, Akbir Khan, Stella Biderman, Sara Hooker, Tim Rocktäschel, and Edward Grefenstette. 2023 · 2023
Later among the works it cites.
Verbosity bias in preference labeling by large language models
Keita Saito, Akifumi Wachi, Koki Wataoka, and Youhei Akimoto. 2023 · 2023
Later among the works it cites.
Towards understanding sycophancy in language models
Original
Mrinank Sharma, Meg Tong, Tomasz Korbak, David Duvenaud, Amanda Askell, Samuel R. Bowman, Newton Cheng, Esin Durmus, Zac Hatfield-Dodds, Scott R. Johnston, Shauna Kravec, Timothy Maxwell, Sam McCandlish, Kamal Ndousse, Oliver Rausch, Nicholas Schiefer, Da Yan, Miranda Zhang, and Ethan Perez. 2023 · 2023
Later among the works it cites.
Humanoid agents: Platform for simulating human-like generative agents
Original
Zhilin Wang, Yu Ying Chiu, and Yu Cheung Chiu. 2023 · 2023
Later among the works it cites.
From word models to world models: Translating from natural language to the probabilistic language of thought
Original
Lionel Wong, Gabriel Grand, Alexander K Lew, Noah D Goodman, Vikash K Mansinghka, Jacob Andreas, and Joshua B Tenenbaum. 2023 · 2023
Later among the works it cites.
Inferring the goals of communicating agents from actions and instructions
Lance Ying, Tan Zhi-Xuan, Vikash Mansinghka, and Joshua B Tenenbaum. 2023 · 2023
Later among the works it cites.
How well can llms negotiate? negotiationarena platform and analysis
Original
Federico Bianchi, Patrick John Chia, Mert Yuksekgonul, Jacopo Tagliabue, Dan Jurafsky, and James Zou. 2024 · 2024
Closest in time.
Under the surface: Tracking the artifactuality of llm-generated data
Original
Debarati Das, Karin De Langis, Anna Martin, Jaehyung Kim, Minhwa Lee, Zae Myung Kim, Shirley Hayati, Risako Owan, Bin Hu, Ritik Parkar, et al. 2024 · 2024
Closest in time.
Mixtral of experts
Original
Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, Gianna Lengyel, Guillaume Bour, Guillaume Lample, Lélio Renard Lavaud, Lucile Saulnier, Marie-Anne Lachaux, Pierre Stock, Sandeep Subramanian, Sophia Yang, Szymon Antoniak, Teven Le Scao, Théophile Gervet, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed. 2024 · 2024
Closest in time.
Self-alignment of large language models via monopolylogue-based social scene simulation
Xianghe Pang, Shuo Tang, Rui Ye, Yuxin Xiong, Bolun Zhang, Yanfeng Wang, and Siheng Chen. 2024 · 2024
Closest in time.
Simulating human strategic behavior: Comparing single and multi-agent llms
Original
Karthik Sreedhar and Lydia Chilton. 2024 · 2024
Closest in time.
Bootstrapping llm-based task-oriented dialogue agents via self-talk
Original
Dennis Ulmer, Elman Mansimov, Kaixiang Lin, Justin Sun, Xibin Gao, and Yi Zhang. 2024 · 2024
Closest in time.
Sotopia: Interactive evaluation for social intelligence in language agents
Xuhui Zhou, Hao Zhu, Leena Mathur, Ruohong Zhang, Zhengyang Qi, Haofei Yu, Louis-Philippe Morency, Yonatan Bisk, Daniel Fried, Graham Neubig, and Maarten Sap. 2024 · 2024
Closest in time.
Can you put it all together: Evaluating conversational agents’ ability to blend skills
Eric Michael Smith, Mary Williamson, Kurt Shuster, Jason Weston, and Y-Lan Boureau. 2020 · 2030
Closest in time.