Fetching the paper…
Reading the bibliography…
Intention is an important and challenging concept in AI.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Freedom and Action
Roderick Chisholm. 1966 · 1966
Earlier work this paper cites.
Intentionality: An essay in the philosophy of mind
John R Searle. 1983 · 1983
Earlier work this paper cites.
Intention is choice with commitment
Philip R. Cohen and Hector J. Levesque. 1990 · 1990
Earlier work this paper cites.
Intention
Gertrude Elizabeth Margaret Anscombe. 2000 · 2000
Earlier work this paper cites.
Multi-agent influence diagrams for representing and solving games
Daphne Koller and Brian Milch. 2003 · 2003
Earlier work this paper cites.
The Basic AI Drives. In Artificial General Intelligence 2008, Proceedings of the First AGI Conference, AGI 2008, March 1-3, 2008, University of Memphis, Memphis, TN, USA (Frontiers in Artificial Intelligence and Applications, Vol. 171) , Pei Wang, Ben Goertzel, and Stan Franklin (Eds.). IOS Press, 483–492
Stephen M. Omohundro. 2008a · 2008
Earlier work this paper cites.
The Basic AI Drives. In Proceedings of the 2008 Conference on Artificial General Intelligence 2008: Proceedings of the First AGI Conference . IOS Press, NLD, 483–492
Stephen M. Omohundro. 2008b · 2008
Earlier work this paper cites.
Intention, Practical Rationality, and Self-Governance
Michael E. Bratman. 2009 · 2009
Earlier work this paper cites.
Causality
Judea Pearl. 2009 · 2009
Earlier work this paper cites.
Formalizing Convergent Instrumental Goals.. In AAAI Workshop: AI, Ethics, and Society
Tsvi Benson-Tilsen and Nate Soares. 2016 · 2016
Earlier work this paper cites.
Actual causality
Joseph Y Halpern. 2016 · 2016
Earlier work this paper cites.
The Definition of Lying and Deception
James Edwin Mahon. 2016 · 2016
Earlier work this paper cites.
Decisions and Dependence in Influence Diagrams. In Proceedings of the Eighth International Conference on Probabilistic Graphical Models (Proceedings of Machine Learning Research, Vol. 52) , Alessandro Antonucci, Giorgio Corani, and Cassio Polpo Campos (Eds.). PMLR, Lugano, Switzerland, 462–473
Ross D. Shachter. 2016 · 2016
Earlier work this paper cites.
Superintelligence
Nick Bostrom. 2017 · 2017
Cited alongside, same era.
BDI Logics for BDI Architectures: Old Problems, New Perspectives
Andreas Herzig, Emiliano Lorini, Laurent Perrussel, and Zhanhao Xiao. 2017 · 2017
Cited alongside, same era.
Towards Formal Definitions of Blameworthiness, Intention, and Moral Responsibility. In Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence, (AAAI-18), the 30th innovative Applications of Artificial Intelligence (IAAI-18), and the 8th AAAI Symposium on Educational Advances in Artificial Intelligence (EAAI-18), New Orleans, Louisiana, USA, February 2-7, 2018 , Sheila A. McIlraith and Kilian Q. Weinberger (Eds.). AAAI Press, 1853–1860
Joseph Y. Halpern and Max Kleiman-Weiner. 2018 · 2018
Cited alongside, same era.
Lies, Bullshit, and Deception in Agent-Oriented Programming Languages. In Proceedings of the 20th International Trust Workshop co-located with AAMAS/IJCAI/ECAI/ICML 2018, Stockholm, Sweden, July 14, 2018 (CEUR Workshop Proceedings, Vol. 2154) , Robin Cohen, Murat Sensoy, and Timothy J. Norman (Eds.). CEUR-WS.org, 50–61
Goal Misgeneralization: Why Correct Specifications Aren’t Enough For Correct Goals
Rohin Shah, Vikrant Varma, Ramana Kumar, Mary Phuong, Victoria Krakovna, Jonathan Uesato, and Zac Kenton. 2022 · 2022
Later among the works it cites.
Talking About Large Language Models
Murray Shanahan. 2022 · 2022
Later among the works it cites.
A Complete Criterion for Value of Information in Soluble Influence Diagrams
Chris van Merwijk, Ryan Carey, and Tom Everitt. 2022 · 2022
Later among the works it cites.
Characterizing Manipulation from AI Systems. In Proceedings of the 3rd ACM Conference on Equity and Access in Algorithms, Mechanisms, and Optimization, EAAMO 2023, Boston, MA, USA, 30 October 2023 - 1 November 2023 . ACM, 6:1–6:13
Micah Carroll, Alan Chan, Henry Ashton, and David Krueger. 2023 · 2023
Later among the works it cites.
On Imperfect Recall in Multi-Agent Influence Diagrams
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alison R. Panisson, Stefan Sarkadi, Peter McBurney, Simon Parsons, and Rafael H. Bordini. 2018 · 2018
Cited alongside, same era.
Quantifying Generalization in Reinforcement Learning. In Proceedings of the 36th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 97) , Kamalika Chaudhuri and Ruslan Salakhutdinov (Eds.). PMLR, 1282–1289
Karl Cobbe, Oleg Klimov, Chris Hesse, Taehoon Kim, and John Schulman. 2019 · 2019
Cited alongside, same era.
Deception in Epistemic Causal Logic
Chiaki Sakama. 2020 · 2020
Cited alongside, same era.
User Tampering in Reinforcement Learning Recommender Systems
Charles Evans and Atoosa Kasirzadeh. 2021 · 2021
Cited alongside, same era.
Agent Incentives: A Causal Perspective. In Thirty-Fifth AAAI Conference on Artificial Intelligence, AAAI 2021, Thirty-Third Conference on Innovative Applications of Artificial Intelligence, IAAI 2021, The Eleventh Symposium on Educational Advances in Artificial Intelligence, EAAI 2021, Virtual Event, February 2-9, 2021 . AAAI Press, 11487–11495
Tom Everitt, Ryan Carey, Eric D. Langlois, Pedro A. Ortega, and Shane Legg. 2021 · 2021
Cited alongside, same era.
Definitions of intent suitable for algorithms
Hal Ashton. 2022 · 2022
Cited alongside, same era.
Goal Misgeneralization in Deep Reinforcement Learning. In International Conference on Machine Learning, ICML 2022, 17-23 July 2022, Baltimore, Maryland, USA (Proceedings of Machine Learning Research, Vol. 162) , Kamalika Chaudhuri, Stefanie Jegelka, Le Song, Csaba Szepesvári, Gang Niu, and Sivan Sabato (Eds.). PMLR, 12004–12019
Lauro Langosco di Langosco, Jack Koch, Lee D. Sharkey, Jacob Pfau, and David Krueger. 2022 · 2022
Cited alongside, same era.
Path-Specific Objectives for Safer Agent Incentives
Sebastian Farquhar et al. 2022 · 2022
Cited alongside, same era.
In-context Learning and Induction Heads
Catherine Olsson, Nelson Elhage, Neel Nanda, Nicholas Joseph, Nova DasSarma, Tom Henighan, Ben Mann, Amanda Askell, Yuntao Bai, Anna Chen, Tom Conerly, Dawn Drain, Deep Ganguli, Zac Hatfield-Dodds, Danny Hernandez, Scott Johnston, Andy Jones, Jackson Kernion, Liane Lovitt, Kamal Ndousse, Dario Amodei, Tom Brown, Jack Clark, Jared Kaplan, Sam McCandlish, and Chris Olah. 2022 · 2022
Cited alongside, same era.
James Fox, Matt MacDermott, Lewis Hammond, Paul Harrenstein, Alessandro Abate, and Michael Wooldridge. 2023 · 2023
Later among the works it cites.
Reasoning about causality in games
Lewis Hammond, James Fox, Tom Everitt, Ryan Carey, Alessandro Abate, and Michael J. Wooldridge. 2023 · 2023
Later among the works it cites.
Intentionality
Pierre Jacob. 2023 · 2023
Later among the works it cites.
Discovering agents
Zachary Kenton, Ramana Kumar, Sebastian Farquhar, Jonathan Richens, Matt MacDermott, and Tom Everitt. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions
Lorenzo Pacchiardi, Alex J. Chan, Sören Mindermann, Ilan Moscovitz, Alexa Y. Pan, Yarin Gal, Owain Evans, and Jan Brauner. 2023 · 2023
Later among the works it cites.
Honesty Is the Best Policy: Defining and Mitigating AI Deception. In Thirty-seventh Conference on Neural Information Processing Systems
Francis Rhys Ward, Francesca Toni, Francesco Belardinelli, and Tom Everitt. 2023 · 2023
Later among the works it cites.
The Rise and Potential of Large Language Model Based Agents: A Survey
Zhiheng Xi, Wenxiang Chen, Xin Guo, Wei He, Yiwen Ding, Boyang Hong, Ming Zhang, Junzhe Wang, Senjie Jin, Enyu Zhou, Rui Zheng, Xiaoran Fan, Xiao Wang, Limao Xiong, Yuhao Zhou, Weiran Wang, Changhao Jiang, Yicheng Zou, Xiangyang Liu, Zhangyue Yin, Shihan Dou, Rongxiang Weng, Wensen Cheng, Qi Zhang, Wenjuan Qin, Yongyan Zheng, Xipeng Qiu, Xuanjing Huang, and Tao Gui. 2023 · 2023
Later among the works it cites.
Joseph Y. Halpern and Evan Piermont. 2024 · 2024
Closest in time.