Fetching the paper…
Reading the bibliography…
One of the main obstacles for developing flexible AI systems is the split between data-based learners and model-based solvers.
Stevan Harnad, ‘The symbol grounding problem’, Physica D: Nonlinear Phenomena
1990
Earlier work this paper cites.
Stephen Muggleton and Luc De Raedt, ‘Inductive logic programming: Theory and methods’, The Journal of Logic Programming
1994
Earlier work this paper cites.
Richard Sutton and Andrew Barto, Introduction to Reinforcement Learning
1998
Earlier work this paper cites.
Roni Khardon, ‘Learning action strategies for planning domains’, Artificial Intelligence
1999
Earlier work this paper cites.
Drew McDermott, ‘The 1998 AI Planning Systems Competition’, Artificial Intelligence Magazine
2000
Earlier work this paper cites.
Ronen I. Brafman and Moshe Tennenholtz, ‘R-max-a general polynomial time algorithm for near-optimal reinforcement learning’, The Journal of Machine Learning Research
2003
Earlier work this paper cites.
Alan Fern, SungWook Yoon, and Robert Givan, ‘Approximate policy iteration with a policy language bias’, in Advances in neural information processing systems
2004
Earlier work this paper cites.
Mario Martín and Hector Geffner, ‘Learning generalized policies from planning examples using concept languages’, Applied Intelligence
2004
Earlier work this paper cites.
Luke S Zettlemoyer, Hanna Pasula, and Leslie Pack Kaelbling, ‘Learning planning rules in noisy stochastic worlds’, in AAAI
2005
Earlier work this paper cites.
Qiang Yang, Kangheng Wu, and Yunfei Jiang, ‘Learning action models from plan examples using weighted max-sat’, Artificial Intelligence
2007
Earlier work this paper cites.
Carlos Diuk, Andre Cohen, and Michael L Littman, ‘An object-oriented representation for efficient reinforcement learning’, in Proc. ICML
2008
Earlier work this paper cites.
Stephen N Cresswell, Thomas L McCluskey, and Margaret M West, ‘Acquiring planning domain models using LOCM’, The Knowledge Engineering Review
2013
Earlier work this paper cites.
Hector Geffner and Blai Bonet, A concise introduction to models and methods for automated planning
2013
Earlier work this paper cites.
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al., ‘Human-level control through deep reinforcement learning’, Nature
2015
Cited alongside, same era.
2017
Cited alongside, same era.
Brenden M Lake, Tomer D Ullman, Joshua B Tenenbaum, and Samuel J Gershman, ‘Building machines that learn and think like people’, Behavioral and Brain Sciences
2017
Cited alongside, same era.
Ankuj Arora, Humbert Fiorino, Damien Pellier, Marc Métivier, and Sylvie Pesty, ‘A review of learning planning action models’, The Knowledge Engineering Review
2018
Cited alongside, same era.
Judea Pearl and Dana Mackenzie, The book of why: the new science of cause and effect
2018
Later among the works it cites.
2018
Later among the works it cites.
Sam Toyer, Felipe Trevizan, Sylvie Thiébaux, and Lexing Xie, ‘Action schema networks: Generalised policies with deep learning’, in AAAI
2018
Later among the works it cites.
Diego Aineto, Sergio Jiménez, Eva Onaindia, and Miquel Ramírez, ‘Model recognition as planning’, in Proc. ICAPS
2019
Closest in time.
Masataro Asai, ‘Unsupervised grounding of plannable first-order logic representation from images’, in Proc. ICAPS
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Masataro Asai and Alex Fukunaga, ‘Classical planning in deep latent space: Bridging the subsymbolic-symbolic boundary’, in AAAI
2018
Cited alongside, same era.
Aniket Nick Bajpai, Sankalp Garg, et al., ‘Transfer of deep reactive policies for mdp planning’, in Advances in Neural Information Processing Systems
2018
Cited alongside, same era.
Adnan Darwiche, ‘Human-level intelligence or animal-like abilities?’, Communications of the ACM
2018
Cited alongside, same era.
Hector Geffner, ‘Model-free, model-based, and general intelligence’, in IJCAI
2018
Cited alongside, same era.
Edward Groshev, Maxwell Goldstein, Aviv Tamar, Siddharth Srivastava, and Pieter Abbeel, ‘Learning generalized reactive policies using deep neural networks’, in Proc. ICAPS
2018
Cited alongside, same era.
Murugeswari Issakkimuthu, Alan Fern, and Prasad Tadepalli, ‘Training deep reactive policies for probabilistic planning problems’, in ICAPS
2018
Cited alongside, same era.
George Konidaris, Leslie Pack Kaelbling, and Tomas Lozano-Perez, ‘From skills to symbols: Learning symbolic representations for abstract high-level planning’, Journal of Artificial Intelligence Research
2018
Cited alongside, same era.
Gary Marcus, ‘Deep learning: A critical appraisal’, arXiv preprint arXiv:1801.00631
2018
Cited alongside, same era.
Gilles Audemard and Laurent Simon, ‘Predicting learnt clauses quality in modern SAT solver’, in Proc. IJCAI
2019
Closest in time.
Blai Bonet, Guillem Francès, and Hector Geffner, ‘Learning features and abstract actions for computing generalized plans’, in Proc. AAAI
2019
Closest in time.
Thiago P Bueno, Leliane N de Barros, Denis D Mauá, and Scott Sanner, ‘Deep reactive policies for planning in stochastic nonlinear domains’, in AAAI
2019
Closest in time.
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, and Yoshua Bengio, ‘Babyai: A platform to study the sample efficiency of grounded language learning’, in ICLR
2019
Closest in time.
Vincent François-Lavet, Yoshua Bengio, Doina Precup, and Joelle Pineau, ‘Combined reinforcement learning via abstract representations’, in Proc. AAAI
2019
Closest in time.
Marta Garnelo and Murray Shanahan, ‘Reconciling deep learning with symbolic artificial intelligence: representing objects and relations’, Current Opinion in Behavioral Sciences
2019
Closest in time.
Patrik Haslum, Nir Lipovetzky, Daniele Magazzeni, and Christian Muise, An Introduction to the Planning Domain Definition Language
2019
Closest in time.