Fetching the paper…
Reading the bibliography…
Robots deployed in many real-world settings need to be able to acquire new skills and solve new tasks over time.
S. Thrun and T. M. Mitchell, “Lifelong robot learning,”
1995
Earlier work this paper cites.
R. S. Sutton, D. Precup, and S. Singh, “Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,”
1999
Earlier work this paper cites.
S. M. LaValle,
2006
Earlier work this paper cites.
L. P. Kaelbling and T. Lozano-Pérez, “Hierarchical task and motion planning in the now,” in
2011
Earlier work this paper cites.
——, “Integrated task and motion planning in belief space,”
2013
Earlier work this paper cites.
J. Butzke, K. Sapkota, K. Prasad, B. MacAllister, and M. Likhachev, “State lattice with controllers: Augmenting lattice-based path planning with controller-based motion primitives,”
2014
Earlier work this paper cites.
M. Hausknecht and P. Stone, “Deep reinforcement learning in parameterized action space,”
2015
Earlier work this paper cites.
E. Ugur and J. Piater, “Bottom-up learning of object categories, action effects and logical rules: From continuous manipulative exploration to symbolic planning,” in
2015
Earlier work this paper cites.
W. Masson, P. Ranchod, and G. Konidaris, “Reinforcement learning with parameterized actions,” in
2016
Earlier work this paper cites.
P. W. Battaglia, R. Pascanu, M. Lai, D. Rezende, and K. Kavukcuoglu, “Interaction networks for learning about objects, relations and physics,”
2016
Earlier work this paper cites.
S.-K. Kim and M. Likhachev, “Parts assembly planning under uncertainty with simulation-aided physical reasoning,” in
2017
Earlier work this paper cites.
A. Srinivas, A. Jabri, P. Abbeel, S. Levine, and C. Finn, “Universal planning networks: Learning generalizable representations for visuomotor control,” in
2018
Earlier work this paper cites.
G. Konidaris, L. P. Kaelbling, and T. Lozano-Perez, “From skills to symbols: Learning symbolic representations for abstract high-level planning,”
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
B. Ames, A. Thackston, and G. Konidaris, “Learning symbolic representations for planning with parameterized skills,”
2018
Earlier work this paper cites.
R. Kartmann, F. Paus, M. Grotz, and T. Asfour, “Extraction of physically plausible support relations to predict and validate manipulation action effects,”
2018
Cited alongside, same era.
J. Liang, V. Makoviychuk, A. Handa, N. Chentanez, M. Macklin, and D. Fox, “Gpu-accelerated robotic simulation for distributed reinforcement learning,”
2018
Cited alongside, same era.
A. Sharma, S. Gu, S. Levine, V. Kumar, and K. Hausman, “Dynamics-aware unsupervised discovery of skills,”
2019
Cited alongside, same era.
S. Nasiriany, V. H. Pong, S. Lin, and S. Levine, “Planning with goal-conditioned policies,”
2019
Cited alongside, same era.
H. Song, J. A. Haustein, W. Yuan, K. Hang, M. Y. Wang, D. Kragic, and J. A. Stork, “Multi-object rearrangement with monte carlo tree search: A case study on planar nonprehensile sorting,”
2019
Cited alongside, same era.
2020
Later among the works it cites.
A. Bagaria, J. Crowley, J. W. N. Lim, and G. Konidaris, “Skill discovery for exploration and planning using deep skill graphs,” 2020
2020
Later among the works it cites.
A. Suárez-Hernández, T. Gaugry, J. Segovia-Aguas, A. Bernardin, C. Torras, M. Marchal, and G. Alenyà, “Leveraging multiple environments for learning and decision making: a dismantling use case,”
2020
Later among the works it cites.
C. R. Garrett, T. Lozano-Pérez, and L. P. Kaelbling, “Pddlstream: Integrating symbolic planners and blackbox samplers via optimistic adaptive planning,” in
2020
Later among the works it cites.
M. Janner, I. Mordatch, and S. Levine, “gamma-models: Generative temporal difference learning for infinite-horizon prediction,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S.-K. Kim, O. Salzman, and M. Likhachev, “Pomhdp: Search-based belief space planning using multiple heuristics,” in
2019
Cited alongside, same era.
M. Eppe, P. D. Nguyen, and S. Wermter, “From semantics to execution: Integrating action planning with reinforcement learning for robotic causal problem-solving,”
2019
Cited alongside, same era.
M. Janner, S. Levine, W. T. Freeman, J. B. Tenenbaum, C. Finn, and J. Wu, “Reasoning about physical interactions with object-oriented prediction and planning,”
2019
Cited alongside, same era.
M. Fey and J. E. Lenssen, “Fast graph representation learning with PyTorch Geometric,” in
2019
Cited alongside, same era.
M. Y. Seker, A. E. Tekden, and E. Ugur, “Deep effect trajectory prediction in robot manipulation,”
2019
Cited alongside, same era.
A. Mandlekar, F. Ramos, B. Boots, S. Savarese, L. Fei-Fei, A. Garg, and D. Fox, “Iris: Implicit reinforcement without interaction at scale for learning control from offline robot manipulation data,” in
2020
Cited alongside, same era.
A. Conkey and T. Hermans, “Planning under uncertainty to goal distributions,”
2020
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
A. Hinrichs, D. Krieg, R. J. Kunsch, and D. Rudolf, “Expected dispersion of uniformly distributed points,”
2020
Later among the works it cites.
K. Lu, A. Grover, P. Abbeel, and I. Mordatch, “Reset-free lifelong learning with skill-space planning,”
2021
Closest in time.
K. Xie, H. Bharadhwaj, D. Hafner, A. Garg, and F. Shkurti, “Skill transfer via partially amortized hierarchical planning,”
2021
Closest in time.
T. Li, R. Calandra, D. Pathak, Y. Tian, F. Meier, and A. Rai, “Planning in learned latent action spaces for generalizable legged locomotion,”
2021
Closest in time.
S. Nasiriany, V. H. Pong, A. Nair, A. Khazatsky, G. Berseth, and S. Levine, “Disco rl: Distribution-conditioned reinforcement learning for general-purpose policies,”
2021
Closest in time.
Z. Wang, C. R. Garrett, L. P. Kaelbling, and T. Lozano-Pérez, “Learning compositional models of robot skills for task and motion planning,”
2021
Closest in time.
2021
Closest in time.
P. Naderian, G. Loaiza-Ganem, H. J. Braviner, A. L. Caterini, J. C. Cresswell, T. Li, and A. Garg, “C-learning: Horizon-aware cumulative accessibility estimation,”
2021
Closest in time.