Fetching the paper…
Reading the bibliography…
People often give instructions whose meaning is ambiguous without further context, expecting that their actions or goals will disambiguate their intentions.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
On the theory of apportionment
William R Thompson. 1935 · 1935
Earlier work this paper cites.
Systematic sampling
Frank Yates. 1948 · 1948
Earlier work this paper cites.
Unified Pragmatic Models for Generating and Following Instructions. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) , Marilyn Walker, Heng Ji, and Amanda Stent (Eds.). Association for Computational Linguistics, New Orleans, Louisiana, 1951–1963
Daniel Fried, Jacob Andreas, and Dan Klein. 2018 · 1963
Earlier work this paper cites.
The traveling-salesman problem and minimum spanning trees: Part II
Michael Held and Richard M Karp. 1971 · 1971
Earlier work this paper cites.
How to do things with words . Vol. 88
John Langshaw Austin. 1975 · 1975
Earlier work this paper cites.
Logic and Conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
Real-time heuristic search
Richard E Korf. 1990 · 1990
Earlier work this paper cites.
Learning to act using real-time dynamic programming
Andrew G Barto, Steven J Bradtke, and Satinder P Singh. 1995 · 1995
Earlier work this paper cites.
Learning policies for partially observable environments: Scaling up
Michael L Littman, Anthony R Cassandra, and Leslie Pack Kaelbling. 1995 · 1995
Earlier work this paper cites.
Planning, learning and coordination in multiagent decision processes. In TARK , Vol. 96. Citeseer, 195–210
Craig Boutilier. 1996 · 1996
Earlier work this paper cites.
PDDL - The Planning Domain Definition Language
Drew McDermott, Malik Ghallab, Adele Howe, Craig Knoblock, Ashwin Ram, Manuela Veloso, Daniel Weld, and David Wilkins. 1998 · 1998
Earlier work this paper cites.
Value-function approximations for partially observable Markov decision processes
Milos Hauskrecht. 2000 · 2000
Earlier work this paper cites.
Rao-Blackwellised particle filtering for dynamic Bayesian networks
Kevin Murphy and Stuart Russell. 2001 · 2001
Earlier work this paper cites.
Teleological reasoning in infancy: The naıve theory of rational action
György Gergely and Gergely Csibra. 2003 · 2003
Earlier work this paper cites.
A cognitive hierarchy model of games
Colin F Camerer, Teck-Hua Ho, and Juin-Kuan Chong. 2004 · 2004
Earlier work this paper cites.
On the graded salience hypothesis
Rachel Giora. 2004 · 2004
Earlier work this paper cites.
The role of salience in processing pragmatic units
Istvan Kecskes et al · 2004
Earlier work this paper cites.
New admissible heuristics for domain-independent planning. In AAAI , Vol. 5. 9–13
Patrik Haslum, Blai Bonet, Héctor Geffner, et al · 2005
Earlier work this paper cites.
Sequential Monte Carlo samplers
Pierre Del Moral, Arnaud Doucet, and Ajay Jasra. 2006 · 2006
Earlier work this paper cites.
Real-Time Adaptive A*. In Proceedings of the fifth international joint conference on Autonomous agents and multiagent systems . 281–288
Sven Koenig and Maxim Likhachev. 2006 · 2006
Earlier work this paper cites.
Bayesian Inverse Reinforcement Learning.. In IJCAI , Vol. 7. 2586–2591
Deepak Ramachandran and Eyal Amir. 2007 · 2007
Earlier work this paper cites.
Shared intentionality
Michael Tomasello and Malinda Carpenter. 2007 · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning.. In AAAI , Vol. 8. Chicago, IL, USA, 1433–1438
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey. 2008 · 2008
Earlier work this paper cites.
Action understanding as inverse planning
Chris L Baker, Rebecca Saxe, and Joshua B Tenenbaum. 2009 · 2009
Earlier work this paper cites.
Monte Carlo sampling methods for approximating interactive POMDPs
Prashant Doshi and Piotr J Gmytrasiewicz. 2009 · 2009
Earlier work this paper cites.
Landmarks, critical paths and abstractions: what’s the difference anyway?. In Proceedings of the International Conference on Automated Planning and Scheduling , Vol. 19. 162–169
Malte Helmert and Carmel Domshlak. 2009 · 2009
Earlier work this paper cites.
Comparing real-time and incremental heuristic search for real-time situated agents
Sven Koenig and Xiaoxun Sun. 2009 · 2009
Earlier work this paper cites.
Probabilistic plan recognition using off-the-shelf classical planners. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 24
Miguel Ramírez and Hector Geffner. 2010 · 2010
Cited alongside, same era.
Bayesian Theory of Mind: Modeling Joint Belief-Desire Attribution. In Proceedings of the Annual Meeting of the Cognitive Science Society, 33 (33)
Chris Baker, Rebecca Saxe, and Joshua Tenenbaum. 2011 · 2011
Cited alongside, same era.
Capir: Collaborative action planning with intention recognition. In Proceedings of the AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment , Vol. 7. 61–66
Truong-Huy Nguyen, David Hsu, Wee-Sun Lee, Tze-Yun Leong, Leslie Kaelbling, Tomas Lozano-Perez, and Andrew Grant. 2011 · 2011
Cited alongside, same era.
Understanding natural language commands for robotic navigation and mobile manipulation. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 25. 1507–1514
Stefanie Tellex, Thomas Kollar, Steven Dickerson, Matthew Walter, Ashis Banerjee, Seth Teller, and Nicholas Roy. 2011 · 2011
Cited alongside, same era.
Planning with uncertain specifications (puns)
Ankit Shah, Shen Li, and Julie Shah. 2020 · 2020
Later among the works it cites.
Bootstrapping an Imagined We for Cooperation.. In CogSci
Ning Tang, Stephanie Stacy, Minglu Zhao, Gabriel Marquez, and Tao Gao. 2020 · 2020
Later among the works it cites.
Online Bayesian Goal Inference for Boundedly Rational Planning Agents
Tan Zhi-Xuan, Jordyn Mann, Tom Silver, Josh Tenenbaum, and Vikash Mansinghka. 2020 · 2020
Later among the works it cites.
Modeling the mistakes of boundedly rational agents within a Bayesian theory of mind
Arwa Alanqary, Gloria Z Lin, Joie Le, Tan Zhi-Xuan, Vikash K Mansinghka, and Joshua B Tenenbaum. 2021 · 2021
Later among the works it cites.
Human-Compatible Artificial Intelligence
Stuart Russell. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
POMCoP: Belief Space Planning for Sidekicks in Cooperative Games.. In AIIDE
Owen Macindoe, Leslie Pack Kaelbling, and Tomás Lozano-Pérez. 2012 · 2012
Cited alongside, same era.
The do-calculus revisited. In Proceedings of the Twenty-Eighth Conference on Uncertainty in Artificial Intelligence . 3–11
Judea Pearl. 2012 · 2012
Cited alongside, same era.
Hierarchical bayesian inverse reinforcement learning
Jaedeug Choi and Kee-Eung Kim. 2014 · 2014
Cited alongside, same era.
Reusing cost-minimal paths for goal-directed navigation in partially known terrains
Carlos Hernández, Tansel Uras, Sven Koenig, Jorge A Baier, Xiaoxun Sun, and Pedro Meseguer. 2015 · 2015
Cited alongside, same era.
Grounding English commands to reward functions. In Robotics: Science and Systems
Shawn Squire, Stefanie Tellex, Dilip Arumugam, and Lei Yang. 2015 · 2015
Cited alongside, same era.
Pruning and preprocessing methods for inventory-aware pathfinding. In 2016 IEEE Conference on Computational Intelligence and Games (CIG) . IEEE, 1–8
Davide Aversa, Sebastian Sardina, and Stavros Vassos. 2016 · 2016
Cited alongside, same era.
Pragmatic language interpretation as probabilistic inference
Noah D Goodman and Michael C Frank. 2016 · 2016
Cited alongside, same era.
Cooperative inverse reinforcement learning. In Advances in neural information processing systems . 3909–3917
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan. 2016 · 2016
Cited alongside, same era.
Richard Shin, Christopher H Lin, Sam Thomson, Charles Chen, Subhro Roy, Emmanouil Antonios Platanios, Adam Pauls, Dan Klein, Jason Eisner, and Benjamin Van Durme. 2021 · 2021
Later among the works it cites.
Modeling communication to coordinate perspectives in cooperation
Stephanie Stacy, Chenfei Li, Minglu Zhao, Yiling Yun, Qingyi Zhao, Max Kleiman-Weiner, and Tao Gao. 2021 · 2021
Later among the works it cites.
Too Many Cooks: Bayesian Inference for Coordinating Multi-Agent Collaboration
Sarah A Wu, Rose E Wang, James A Evans, Joshua B Tenenbaum, David C Parkes, and Max Kleiman-Weiner. 2021 · 2021
Later among the works it cites.
Do as I can, not as I say: Grounding language in robotic affordances
Michael Ahn, Anthony Brohan, Noah Brown, Yevgen Chebotar, Omar Cortes, Byron David, Chelsea Finn, Chuyuan Fu, Keerthana Gopalakrishnan, Karol Hausman, et al · 2022
Later among the works it cites.
Inferring rewards from language in context
Jessy Lin, Daniel Fried, Dan Klein, and Anca Dragan. 2022 · 2022
Later among the works it cites.
Lang2LTL: Translating Natural Language Commands to Temporal Specification with Large Language Models. In Workshop on Language and Robotics at CoRL 2022
Jason Xinyu Liu, Ziyi Yang, Benjamin Schornstein, Sam Liang, Ifrah Idrees, Stefanie Tellex, and Ankit Shah. 2022 · 2022
Later among the works it cites.
Exploring an Imagined “We” in human collective hunting: Joint commitment within shared intentionality. In Proceedings of the annual meeting of the cognitive science society , Vol. 44
Ning Tang, Siyi Gong, Minglu Zhao, Chenya Gu, Jifan Zhou, Mowei Shen, and Tao Gao. 2022 · 2022
Later among the works it cites.
HandMeThat: Human-Robot Communication in Physical and Social Environments
Yanming Wan, Jiayuan Mao, and Josh Tenenbaum. 2022 · 2022
Later among the works it cites.
Solving the Baby Intuition Benchmark with a hierarchically Bayesian theory-of-mind
Tan Zhi-Xuan, Nishad Gothoskar, Falk Pollok, Dan Gutfreund, Joshua B Tenenbaum, and Vikash K Mansinghka. 2022 · 2022
Later among the works it cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Later among the works it cites.
Can LLMs Generate Random Numbers? Evaluating LLM Sampling in Controlled Domains. In ICML 2023 Workshop: Sampling and Optimization in Discrete Space
Aspen K Hopkins, Alex Renda, and Michael Carbin. 2023 · 2023
Later among the works it cites.
Understanding the effects of RLHF on LLM generalisation and diversity
Robert Kirk, Ishita Mediratta, Christoforos Nalmpantis, Jelena Luketina, Eric Hambro, Edward Grefenstette, and Roberta Raileanu. 2023 · 2023
Later among the works it cites.
Reward Design with Language Models. In The Eleventh International Conference on Learning Representations
Minae Kwon, Sang Michael Xie, Kalesha Bullard, and Dorsa Sadigh. 2023 · 2023
Later among the works it cites.
Sequential Monte Carlo Steering of Large Language Models using Probabilistic Programs
Alexander K Lew, Tan Zhi-Xuan, Gabriel Grand, and Vikash K Mansinghka. 2023b · 2023
Later among the works it cites.
Minding Language Models’ (Lack of) Theory of Mind: A Plug-and-Play Multi-Character Belief Tracker. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Anna Rogers, Jordan Boyd-Graber, and Naoaki Okazaki (Eds.). Association for Computational Linguistics, Toronto, Canada, 13960–13980
Melanie Sclar, Sachin Kumar, Peter West, Alane Suhr, Yejin Choi, and Yulia Tsvetkov. 2023 · 2023
Later among the works it cites.
Progprompt: Generating situated robot task plans using large language models. In 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 11523–11530
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg. 2023 · 2023
Later among the works it cites.
Reconciling truthfulness and relevance as epistemic and decision-theoretic utility
Theodore R Sumers, Mark K Ho, Thomas L Griffiths, and Robert D Hawkins. 2023 · 2023
Later among the works it cites.
Efficient Guided Generation for LLMs
Brandon T Willard and Rémi Louf. 2023 · 2023
Later among the works it cites.
Lance Ying, Katherine M Collins, Megan Wei, Cedegao E Zhang, Tan Zhi-Xuan, Adrian Weller, Joshua B Tenenbaum, and Lionel Wong. 2023a · 2023
Later among the works it cites.
Language to Rewards for Robotic Skill Synthesis
Wenhao Yu, Nimrod Gileadi, Chuyuan Fu, Sean Kirmani, Kuang-Huei Lee, Montse Gonzalez Arenas, Hao-Tien Lewis Chiang, Tom Erez, Leonard Hasenclever, Jan Humplik, et al · 2023
Later among the works it cites.
Visual cognition in multimodal large language models
Luca M. Schulze Buschoff, Elif Akata, Matthias Bethge, and Eric Schulz. 2024 · 2024
Closest in time.
MMToM-QA: Multimodal Theory of Mind Question Answering
Chuanyang Jin, Yutong Wu, Jing Cao, Jiannan Xiang, Yen-Ling Kuo, Zhiting Hu, Tomer Ullman, Antonio Torralba, Joshua B Tenenbaum, and Tianmin Shu. 2024 · 2024
Closest in time.