Fetching the paper…
Reading the bibliography…
To achieve human-like common sense about everyday life, machine learning systems must understand and reason about the goals, preferences, and actions of other agents in the environment.
An experimental study of apparent behavior
Heider, F. and Simmel, M. (1944) · 1944
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
Premack, D. and Woodruff, G. (1978) · 1978
Earlier work this paper cites.
Object permanence in five-month-old infants
Baillargeon, R., Spelke, E. S., and Wasserman, S. (1985) · 1985
Earlier work this paper cites.
Does the autistic child have a “theory of mind”?
Baron-Cohen, S., Leslie, A. M., and Frith, U. (1985) · 1985
Earlier work this paper cites.
Object permanence in 3 1 / 2 1/2 -and 4 1 / 2 1/2 -month-old infants
Baillargeon, R. (1987) · 1987
Earlier work this paper cites.
The development of young infants’ intuitions about support
Baillargeon, R., Needham, A., and DeVos, J. (1992) · 1992
Earlier work this paper cites.
Origins of knowledge
Spelke, E. S., Breinlinger, K., Macomber, J., and Jacobson, K. (1992) · 1992
Earlier work this paper cites.
Taking the intentional stance at 12 months of age
Gergely, G., Nádasdy, Z., Csibra, G., and Bíró, S. (1995) · 1995
Earlier work this paper cites.
Teleological reasoning in infancy: The infant’s naive theory of rational action: A reply to premack and premack
Gergely, G. and Csibra, G. (1997) · 1997
Earlier work this paper cites.
Early reasoning about desires: evidence from 14-and 18-month-olds
Repacholi, B. M. and Gopnik, A. (1997) · 1997
Earlier work this paper cites.
Infants selectively encode the goal object of an actor’s reach
Woodward, A. L. (1998) · 1998
Earlier work this paper cites.
Infants’ ability to distinguish between purposeful and non-purposeful behaviors
Woodward, A. L. (1999) · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Ng, A. Y., Russell, S. J., et al. (2000) · 2000
Earlier work this paper cites.
Twelve-month-old infants interpret action in context
Woodward, A. L. and Sommerville, J. A. (2000) · 2000
Earlier work this paper cites.
Rational imitation in preverbal infants
Gergely, G., Bekkering, H., and Király, I. (2002) · 2002
Earlier work this paper cites.
Teleological reasoning in infancy: The naıve theory of rational action
Gergely, G. and Csibra, G. (2003) · 2003
Earlier work this paper cites.
Attribution of dispositional states by 12-month-olds
Kuhlmeier, V., Wynn, K., and Bloom, P. (2003) · 2003
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y. (2004) · 2004
Earlier work this paper cites.
Infants’ attribution of a goal to a morphologically unfamiliar agent
Shimizu, Y. A. and Johnson, S. C. (2004) · 2004
Earlier work this paper cites.
Do infants apply the principle of rational action to human agents?
Sodian, B., Schoeppner, B., and Metz, U. (2004) · 2004
Earlier work this paper cites.
Twelve-and 18-month-olds copy actions in terms of goals
Carpenter, M., Call, J., and Tomasello, M. (2005) · 2005
Earlier work this paper cites.
Offline reinforcement learning: Tutorial, review, and perspectives on open problems
Levine, S., Kumar, A., Tucker, G., and Fu, J. (2020) · 2005
Earlier work this paper cites.
Can a self-propelled box have a goal? psychological reasoning in 5-month-old infants
Luo, Y. and Baillargeon, R. (2005) · 2005
Earlier work this paper cites.
Infants’ understanding of object-directed action
Phillips, A. T. and Wellman, H. M. (2005) · 2005
Earlier work this paper cites.
Pulling out the intentional structure of action: the relation between action processing and action production in infancy
Sommerville, J. A. and Woodward, A. L. (2005) · 2005
Cited alongside, same era.
Can infants attribute to an agent a disposition to perform a particular action?
Song, H.-j., Baillargeon, R., and Fisher, C. (2005) · 2005
Cited alongside, same era.
Five-month-old infants know humans are solid, like inanimate objects
Saxe, R., Tzelnic, T., and Carey, S. (2006) · 2006
Cited alongside, same era.
Infants track action goals within and across agents
Buresh, J. S. and Woodward, A. L. (2007) · 2007
Cited alongside, same era.
Imitating step by step: A detailed analysis of 9-to 15-month-olds’ reproduction of a three-step action sequence
Elsner, B., Hauf, P., and Aschersleben, G. (2007) · 2007
Cited alongside, same era.
Psychological reasoning in infancy
Baillargeon, R., Scott, R. M., and Bian, L. (2016) · 2016
Later among the works it cites.
Generative adversarial imitation learning
Ho, J. and Ermon, S. (2016) · 2016
Later among the works it cites.
Core knowledge and conceptual change
Spelke, E. S. (2016) · 2016
Later among the works it cites.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Baker, C. L., Jara-Ettinger, J., Saxe, R., and Tenenbaum, J. B. (2017) · 2017
Later among the works it cites.
Building machines that learn and think like people
Lake, B. M., Ullman, T. D., Tenenbaum, J. B., and Gershman, S. J. (2017) · 2017
Later among the works it cites.
Six-month-old infants expect agents to minimize the cost of their actions
Liu, S. and Spelke, E. S. (2017) · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Saxe, R., Tzelnic, T., and Carey, S. (2007) · 2007
Cited alongside, same era.
Infants attribute goals to biomechanically impossible actions
Southgate, V., Johnson, M., and Csibra, G. (2008) · 2008
Cited alongside, same era.
Babies and brains: habituation in infant cognition and functional neuroimaging
Turk-Browne, N. B., Scholl, B. J., and Chun, M. M. (2008) · 2008
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
Ziebart, B. D., Maas, A. L., Bagnell, J. A., and Dey, A. K. (2008) · 2008
Cited alongside, same era.
Action understanding as inverse planning
Baker, C. L., Saxe, R., and Tenenbaum, J. B. (2009) · 2009
Cited alongside, same era.
Young infants’ reasoning about physical events involving inert and self-propelled objects
Luo, Y., Kaufman, L., and Baillargeon, R. (2009) · 2009
Cited alongside, same era.
Help or hinder: Bayesian models of social goal inference
Ullman, T., Baker, C., Macindoe, O., Evans, O., Goodman, N., and Tenenbaum, J. B. (2009) · 2009
Cited alongside, same era.
Ten-month-old infants infer the value of goals from the costs of actions
Liu, S., Ullman, T. D., Tenenbaum, J. B., and Spelke, E. S. (2017) · 2017
Later among the works it cites.
Maximum a posteriori policy optimisation
Abdolmaleki, A., Springenberg, J. T., Tassa, Y., Munos, R., Heess, N., and Riedmiller, M. (2018) · 2018
Later among the works it cites.
Autonomous agents modelling other agents: A comprehensive survey and open problems
Albrecht, S. V. and Stone, P. (2018) · 2018
Later among the works it cites.
Machine theory of mind
Rabinowitz, N., Perbet, F., Song, F., Zhang, C., Eslami, S. M. A., and Botvinick, M. (2018) · 2018
Later among the works it cites.
Modeling others using oneself in multi-agent reinforcement learning
Raileanu, R., Denton, E., Szlam, A., and Fergus, R. (2018) · 2018
Later among the works it cites.
Intphys: A framework and benchmark for visual intuitive physics reasoning
Riochet, R., Castro, M. Y., Bernard, M., Lerer, A., Fergus, R., Izard, V., and Dupoux, E. (2018) · 2018
Later among the works it cites.
Theory of mind as inverse reinforcement learning
Jara-Ettinger, J. (2019) · 2019
Later among the works it cites.
Ai2-thor: An interactive 3d environment for visual ai
Kolve, E., Mottaghi, R., Han, W., VanderBilt, E., Weihs, L., Herrasti, A., Gordon, D., Zhu, Y., Gupta, A., and Farhadi, A. (2019) · 2019
Later among the works it cites.
Origins of the concepts cause, cost, and goal in prereaching infants
Liu, S., Brooks, N. B., and Spelke, E. S. (2019) · 2019
Later among the works it cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
Rakelly, K., Zhou, A., Finn, C., Levine, S., and Quillen, D. (2019) · 2019
Later among the works it cites.
Modeling expectation violation in intuitive physics with coarse probabilistic object representations
Smith, K., Mei, L., Yao, S., Wu, J., Spelke, E., Tenenbaum, J., and Ullman, T. (2019) · 2019
Later among the works it cites.
Learning a prior over intent via meta-inverse reinforcement learning
Xu, K., Ratner, E., Dragan, A., Levine, S., and Finn, C. (2019) · 2019
Later among the works it cites.
Meta-inverse reinforcement learning with probabilistic context variables
Yu, L., Yu, T., Finn, C., and Ermon, S. (2019) · 2019
Later among the works it cites.
Efficiency as a principle for social preferences in infancy
Colomer, M., Bas, J., and Sebastian-Galles, N. (2020) · 2020
Later among the works it cites.
Keep doing what worked: Behavior modelling priors for offline reinforcement learning
Siegel, N., Springenberg, J. T., Berkenkamp, F., Abdolmaleki, A., Neunert, M., Lampe, T., Hafner, R., Heess, N., and Riedmiller, M. (2020) · 2020
Later among the works it cites.
AGENT: A Benchmark for Core Psychological Reasoning
Shu, T., Bhandwaldar, A., Gan, C., Smith, K., Liu, S., Gutfreund, D., Spelke, E., Tenenbaum, J. B., and Ullman, T. D. (2021) · 2021
Closest in time.
Decoupling representation learning from reinforcement learning
Stooke, A., Lee, K., Abbeel, P., and Laskin, M. (2021) · 2021
Closest in time.