Fetching the paper…
Reading the bibliography…
Decision-making is challenging in robotics environments with continuous object-centric states, continuous actions, long horizons, and sparse feedback.
STRIPS: A new approach to the application of theorem proving to problem solving
R. E. Fikes and N. J. Nilsson · 1971
Earlier work this paper cites.
Feudal reinforcement learning
P. Dayan and G. E. Hinton · 1992
Earlier work this paper cites.
Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning
R. S. Sutton, D. Precup, and S. Singh · 1999
Earlier work this paper cites.
FF: The fast-forward planning system
J. Hoffmann · 2001
Earlier work this paper cites.
Automatic discovery of subgoals in reinforcement learning using diverse density
A. McGovern and A. G. Barto · 2001
Earlier work this paper cites.
Using abstract models of behaviours to automatically generate reinforcement learning hierarchies
M. R. Ryan · 2002
Earlier work this paper cites.
aSyMov: a planner that deals with intricate symbolic and geometric problems
F. Gravot, S. Cambon, and R. Alami · 2005
Earlier work this paper cites.
Combining reinforcement learning with symbolic planning
M. Grounds and D. Kudenko · 2005
Earlier work this paper cites.
Towards a unified theory of state abstraction for MDPs
L. Li, T. J. Walsh, and M. L. Littman · 2006
Earlier work this paper cites.
The fast downward planning system
M. Helmert · 2006
Earlier work this paper cites.
Angelic semantics for high-level actions
B. Marthi, S. J. Russell, and J. A. Wolfe · 2007
Earlier work this paper cites.
Learning symbolic models of stochastic domains
H. M. Pasula, L. S. Zettlemoyer, and L. P. Kaelbling · 2007
Earlier work this paper cites.
Learning action models from plan examples using weighted max-sat
Q. Yang, K. Wu, and Y. Jiang · 2007
Earlier work this paper cites.
Landmarks, critical paths and abstractions: what’s the difference anyway?
M. Helmert and C. Domshlak · 2009
Earlier work this paper cites.
Skill discovery in continuous reinforcement learning domains using skill chaining
G. Konidaris and A. Barto · 2009
Earlier work this paper cites.
Learning complex action models with quantifiers and logical implications
H. H. Zhuo, Q. Yang, D. H. Hu, and L. Li · 2010
Earlier work this paper cites.
Learning and generalization of complex tasks from unstructured demonstrations
S. Niekum, S. Osentoski, G. Konidaris, and A. G. Barto · 2012
Earlier work this paper cites.
B. Da Silva, G. Konidaris, and A. Barto · 2012
Earlier work this paper cites.
Learning grounded relational symbols from continuous data for abstract reasoning
N. Jetchev, T. Lang, and M. Toussaint · 2013
Earlier work this paper cites.
Refining incomplete planning domain models through plan traces
H. H. Zhuo, T. Nguyen, and S. Kambhampati · 2013
Earlier work this paper cites.
Combined task and motion planning through an extensible planner-independent interface layer
S. Srivastava, E. Fang, L. Riano, R. Chitnis, S. Russell, and P. Abbeel · 2014
Earlier work this paper cites.
New algorithms for the top-k planning problem
A. Riabov, S. Sohrabi, and O. Udrea · 2014
Earlier work this paper cites.
Pac-inspired option discovery in lifelong reinforcement learning
E. Brunskill and L. Li · 2014
Earlier work this paper cites.
Learning parameterized motor skills on a humanoid robot
B. C. Da Silva, G. Baldassarre, G. Konidaris, and A. Barto · 2014
Earlier work this paper cites.
Logic-geometric programming: An optimization-based approach to combined task and motion planning
M. Toussaint · 2015
Earlier work this paper cites.
Nonparametric bayesian reward segmentation for skill discovery using inverse reinforcement learning
P. Ranchod, B. Rosman, and G. Konidaris · 2015
Earlier work this paper cites.
Bottom-up learning of object categories, action effects and logical rules: From continuous manipulative exploration to symbolic planning
E. Ugur and J. Piater · 2015
Earlier work this paper cites.
Do what I want, not what I did: Imitation of skills by planning sequences of actions
C. Paxton, F. Jonathan, M. Kobilarov, and G. D. Hager · 2016
Earlier work this paper cites.
Near optimal behavior via approximate state abstraction
D. Abel, D. Hershkowitz, and M. Littman · 2016
Earlier work this paper cites.
Incremental task and motion planning: A constraint-based approach
N. T. Dantam, Z. K. Kingston, S. Chaudhuri, and L. E. Kavraki · 2016
Earlier work this paper cites.
Guided search for task and motion plans using learned heuristics
R. Chitnis, D. Hadfield-Menell, A. Gupta, S. Srivastava, E. Groshev, C. Lin, and P. Abbeel · 2016
Earlier work this paper cites.
K. Gregor, D. J. Rezende, and D. Wierstra · 2016
Cited alongside, same era.
Probabilistic inference for determining options in reinforcement learning
C. Daniel, H. Van Hoof, J. Peters, and G. Neumann · 2016
Cited alongside, same era.
Hindsight experience replay
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, O. Pieter Abbeel, and W. Zaremba · 2017
Cited alongside, same era.
The option-critic architecture
P.-L. Bacon, J. Harb, and D. Precup · 2017
Cited alongside, same era.
Independently controllable features
V. Thomas, J. Pondard, E. Bengio, M. Sarfati, P. Beaudoin, M.-J. Meurs, J. Pineau, D. Precup, and Y. Bengio · 2017
Cited alongside, same era.
Multi-level discovery of deep options
Accelerating reinforcement learning with learned skill priors
K. Pertsch, Y. Lee, and J. J. Lim · 2020
Later among the works it cites.
Learning portable representations for high-level planning
S. James, B. Rosman, and G. Konidaris · 2020
Later among the works it cites.
Integrated task and motion planning
C. R. Garrett, R. Chitnis, R. Holladay, B. Kim, T. Silver, L. P. Kaelbling, and T. Lozano-Perez · 2021
Later among the works it cites.
RePReL: Integrating relational planning and reinforcement learning for effective abstraction
H. Kokel, A. Manoharan, S. Natarajan, B. Ravindran, and P. Tadepalli · 2021
Later among the works it cites.
Learning symbolic operators for task and motion planning
T. Silver, R. Chitnis, J. Tenenbaum, L. P. Kaelbling, and T. Lozano-Pérez · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Fox, S. Krishnan, I. Stoica, and K. Goldberg · 2017
Cited alongside, same era.
Feudal networks for hierarchical reinforcement learning
A. S. Vezhnevets, S. Osindero, T. Schaul, N. Heess, M. Jaderberg, D. Silver, and K. Kavukcuoglu · 2017
Cited alongside, same era.
From skills to symbols: Learning symbolic representations for abstract high-level planning
G. Konidaris, L. P. Kaelbling, and T. Lozano-Perez · 2018
Cited alongside, same era.
Data-efficient hierarchical reinforcement learning
O. Nachum, S. Gu, H. Lee, and S. Levine · 2018
Cited alongside, same era.
A novel iterative approach to top-k planning
M. Katz, S. Sohrabi, O. Udrea, and D. Winterer · 2018
Cited alongside, same era.
Differentiable physics and stable modes for tool-use and manipulation planning
M. A. Toussaint, K. R. Allen, K. A. Smith, and J. B. Tenenbaum · 2018
Cited alongside, same era.
A review of learning planning action models
A. Arora, H. Fiorino, D. Pellier, M. Métivier, and S. Pesty · 2018
Cited alongside, same era.
M. Asai and C. Muise · 2021
Later among the works it cites.
Learning a symbolic planning domain through the interaction with continuous environments
E. Umili, E. Antonioni, F. Riccio, R. Capobianco, D. Nardi, and G. De Giacomo · 2021
Later among the works it cites.
Learning geometric reasoning and control for long-horizon tasks from visual input
D. Driess, J.-S. Ha, R. Tedrake, and M. Toussaint · 2021
Later among the works it cites.
Extended tree search for robot task and motion planning
T. Ren, G. Chalvatzaki, and J. Peters · 2021
Later among the works it cites.
Parrot: Data-driven behavioral priors for reinforcement learning
A. Singh, H. Liu, G. Zhou, A. Yu, N. Rhinehart, and S. Levine · 2021
Later among the works it cites.
OPAL: Offline primitive discovery for accelerating offline reinforcement learning
A. Ajay, A. Kumar, P. Agrawal, S. Levine, and O. Nachum · 2021
Later among the works it cites.
SPOTTER: Extending symbolic planning operators through targeted reinforcement learning
V. Sarathy, D. Kasenberg, S. Goel, J. Sinapov, and M. Scheutz · 2021
Later among the works it cites.
Skid raw: Skill discovery from raw trajectories
D. Tanneberg, K. Ploeger, E. Rueckert, and J. Peters · 2021
Later among the works it cites.
Skill induction and planning with latent language
P. Sharma, A. Torralba, and J. Andreas · 2021
Later among the works it cites.
Creativity of AI: Automatic symbolic option discovery for facilitating deep reinforcement learning
M. Jin, Z. Ma, K. Jin, H. H. Zhuo, C. Chen, and C. Yu · 2021
Later among the works it cites.
Learning compositional models of robot skills for task and motion planning
Z. Wang, C. R. Garrett, L. P. Kaelbling, and T. Lozano-Pérez · 2021
Later among the works it cites.
Temporal logic imitation: Learning plan-satisficing motion policies from demonstrations, 2022
Y. Wang, N. Figueroa, S. Li, A. Shah, and J. Shah · 2022
Closest in time.
Learning neuro-symbolic relational transition models for bilevel planning
R. Chitnis, T. Silver, J. B. Tenenbaum, T. Lozano-Perez, and L. P. Kaelbling · 2022
Closest in time.
Leveraging approximate symbolic models for reinforcement learning via skill diversity
L. Guan, S. Sreedharan, and S. Kambhampati · 2022
Closest in time.
Generalizable task planning through representation pretraining
C. Wang, D. Xu, and L. Fei-Fei · 2022
Closest in time.
Inventing relational state and action abstractions for effective and efficient bilevel planning
T. Silver, R. Chitnis, N. Kumar, W. McClinton, T. Lozano-Perez, L. P. Kaelbling, and J. Tenenbaum · 2022
Closest in time.
Bottom-up skill discovery from unsegmented demonstrations for long-horizon robot manipulation
Y. Zhu, P. Stone, and Y. Zhu · 2022
Closest in time.
Implicit behavioral cloning
P. Florence, C. Lynch, A. Zeng, O. A. Ramirez, A. Wahid, L. Downs, A. Wong, J. Lee, I. Mordatch, and J. Tompson · 2022
Closest in time.
Learning temporally extended skills in continuous domains as symbolic actions for planning
J. Achterhold, M. Krimmel, and J. Stueckler · 2022
Closest in time.
Do as i can and not as i say: Grounding language in robotic affordances
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, D. Ho, J. Hsu, J. Ibarz, B. Ichter, A. Irpan, E. Jang, R. J. Ruano, K. Jeffrey, S. Jesmonth, N. Joshi, R. Julian, D. Kalashnikov, Y. Kuang, K.-H. Lee, S. Levine, Y. Lu, L. Luu, C. Parada, P. Pastor, J. Quiambao, K. Rao, J. Rettinghouse, D. Reyes, P. Sermanet, N. Sievers, C. Tan, A. Toshev, V. Vanhoucke, F. Xia, T. Xiao, P. Xu, S. Xu, and M. Yan · 2022
Closest in time.
Pre-trained language models for interactive decision-making
S. Li, X. Puig, Y. Du, C. Wang, E. Akyurek, A. Torralba, J. Andreas, and I. Mordatch · 2022
Closest in time.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch · 2022
Closest in time.
Discovering user-interpretable capabilities of black-box planning agents
P. Verma, S. R. Marpally, and S. Srivastava · 2022
Closest in time.
Structured deep generative models for sampling on constraint manifolds in sequential manipulation
J. Ortiz-Haro, J.-S. Ha, D. Driess, and M. Toussaint · 2022
Closest in time.
Discovering state and action abstractions for generalized task and motion planning
A. Curtis, T. Silver, J. B. Tenenbaum, T. Lozano-Perez, and L. P. Kaelbling · 2022
Closest in time.