Fetching the paper…
Reading the bibliography…
One promising approach towards effective robot decision making in complex, long-horizon tasks is to sequence together parameterized skills.
Online replanning in belief space for partially observable task and motion problems
Caelan Reed Garrett, Chris Paxton, Tomás Lozano-Pérez, Leslie Pack Kaelbling, and Dieter Fox · 1911
Earlier work this paper cites.
Bandit problems. sequential allocation of experiments
Donald A. Berry and Bert Fristedt · 1987
Earlier work this paper cites.
A lifelong learning perspective for mobile robot control
Sebastian Thrun · 1995
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
Peter Auer · 2002
Earlier work this paper cites.
Near-optimal reinforcement learning in polynomial time
Michael Kearns and Satinder Singh · 2002
Earlier work this paper cites.
Pddl2. 1: An extension to pddl for expressing temporal planning domains
Maria Fox and Derek Long · 2003
Earlier work this paper cites.
The fast downward planning system
Malte Helmert · 2006
Earlier work this paper cites.
Learning compositional models of robot skills for task and motion planning
Zi Wang, Caelan Reed Garrett, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2006
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
Carlos Diuk, Andre Cohen, and Michael L Littman · 2008
Earlier work this paper cites.
Pure exploration in multi-armed bandits problems
Sébastien Bubeck, Rémi Munos, and Gilles Stoltz · 2009
Earlier work this paper cites.
A contextual-bandit approach to personalized news article recommendation
Lihong Li, Wei Chu, John Langford, and Robert E Schapire · 2010
Earlier work this paper cites.
Accelerating reinforcement learning with learned skill priors
Karl Pertsch, Youngwoon Lee, and Joseph Lim · 2010
Earlier work this paper cites.
Competence progress intrinsic motivation
Andrew Stout and Andrew G Barto · 2010
Earlier work this paper cites.
From theories to queries: Active learning in practice
Burr Settles · 2011
Earlier work this paper cites.
Optimistic bayesian sampling in contextual-bandit problems
Benedict C May, Nathan Korda, Anthony Lee, and David S Leslie · 2012
Earlier work this paper cites.
Active learning of inverse models with intrinsically motivated goal exploration in robots
Adrien Baranes and Pierre-Yves Oudeyer · 2013
Earlier work this paper cites.
Active learning of parameterized skills
Bruno Da Silva, George Konidaris, and Andrew Barto · 2014
Earlier work this paper cites.
Active reward learning
Christian Daniel, Malte Viering, Jan Metz, Oliver Kroemer, and Jan Peters · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization, 2014
Diederik P. Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Combined task and motion planning through an extensible planner-independent interface layer
Siddharth Srivastava, Eugene Fang, Lorenzo Riano, Rohan Chitnis, Stuart Russell, and Pieter Abbeel · 2014
Earlier work this paper cites.
Simple regret for infinitely many armed bandits
Alexandra Carpentier and Michal Valko · 2015
Earlier work this paper cites.
Guided search for task and motion plans using learned heuristics
Rohan Chitnis, Dylan Hadfield-Menell, Abhishek Gupta, Siddharth Srivastava, Edward Groshev, Christopher Lin, and Pieter Abbeel · 2016
Cited alongside, same era.
Deep reinforcement learning in parameterized action space
Matthew Hausknecht and Peter Stone · 2016
Cited alongside, same era.
Reinforcement learning with parameterized actions
Warwick Masson, Pravesh Ranchod, and George Konidaris · 2016
Cited alongside, same era.
Active exploration for learning symbolic representations
Garrett Andersen and George Konidaris · 2017
Cited alongside, same era.
Learning symbolic representations for planning with parameterized skills
Barrett Ames, Allison Thackston, and George Konidaris · 2018
Cited alongside, same era.
Guiding search in continuous state-action spaces by learning an action sampler from off-target search experience
Learning symbolic operators for task and motion planning
Tom Silver, Rohan Chitnis, Joshua Tenenbaum, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2021
Later among the works it cites.
Parrot: Data-driven behavioral priors for reinforcement learning
Avi Singh, Huihan Liu, Gaoyue Zhou, Albert Yu, Nicholas Rhinehart, and Sergey Levine · 2021
Later among the works it cites.
Planning to practice: Efficient online fine-tuning by composing goals in latent space
Kuan Fang, Patrick Yin, Ashvin Nair, and Sergey Levine · 2022
Later among the works it cites.
Augmenting reinforcement learning with behavior primitives for diverse manipulation tasks
Soroush Nasiriany, Huihan Liu, and Yuke Zhu · 2022
Later among the works it cites.
Learning neuro-symbolic skills for bilevel planning
Tom Silver, Ashay Athalye, Joshua B. Tenenbaum, Tomás Lozano-Pérez, and Leslie Pack Kaelbling · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Beomjoon Kim, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2018
Cited alongside, same era.
From skills to symbols: Learning symbolic representations for abstract high-level planning
George Konidaris, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2018
Cited alongside, same era.
Curious: intrinsically motivated modular multi-goal reinforcement learning
Cédric Colas, Pierre Fournier, Mohamed Chetouani, Olivier Sigaud, and Pierre-Yves Oudeyer · 2019
Cited alongside, same era.
Self-supervised exploration via disagreement
Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta · 2019
Cited alongside, same era.
Intrinsically motivated open-ended learning in autonomous robots, 2020
Vieri Giuliano Santucci, Pierre-Yves Oudeyer, Andrew Barto, and Gianluca Baldassarre · 2019
Cited alongside, same era.
Learning feasibility for task and motion planning in tabletop environments
Andrew M Wells, Neil T Dantam, Anshumali Shrivastava, and Lydia E Kavraki · 2019
Cited alongside, same era.
Skill-based curiosity for intrinsically motivated reinforcement learning
Nicolas Bougie and Ryutaro Ichise · 2020
Cited alongside, same era.
Later among the works it cites.
Accelerating integrated task and motion planning with neural feasibility checking
Lei Xu, Tianyu Ren, Georgia Chalvatzaki, and Jan Peters · 2022
Later among the works it cites.
Detecting twenty-thousand classes using image-level supervision
Xingyi Zhou, Rohit Girdhar, Armand Joulin, Philipp Krähenbühl, and Ishan Misra · 2022
Later among the works it cites.
Diffusion policy: Visuomotor policy learning via action diffusion
Cheng Chi, Siyuan Feng, Yilun Du, Zhenjia Xu, Eric Cousineau, Benjamin Burchfiel, and Shuran Song · 2023
Later among the works it cites.
Kuan Fang, Toki Migimatsu, Ajay Mandlekar, Li Fei-Fei, and Jeannette Bohg · 2023
Later among the works it cites.
Bootstrapped autonomous practicing via multi-task reinforcement learning
Abhishek Gupta, Corey Lynch, Brandon Kinman, Garrett Peake, Sergey Levine, and Karol Hausman · 2023
Later among the works it cites.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Embodied active learning of relational state abstractions for bilevel planning
Amber Li and Tom Silver · 2023
Later among the works it cites.
Eureka: Human-level reward design via coding large language models
Yecheng Jason Ma, William Liang, Guanzhi Wang, De-An Huang, Osbert Bastani, Dinesh Jayaraman, Yuke Zhu, Linxi Fan, and Anima Anandkumar · 2023
Later among the works it cites.
Embodied lifelong learning for task and motion planning
Jorge Mendez-Mendez, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2023
Later among the works it cites.
Generative skill chaining: Long-horizon skill planning with diffusion models
Utkarsh Aashu Mishra, Shangjie Xue, Yongxin Chen, and Danfei Xu · 2023
Later among the works it cites.
Predicate invention for bilevel planning
Tom Silver, Rohan Chitnis, Nishanth Kumar, Willie McClinton, Tomás Lozano-Pérez, Leslie Kaelbling, and Joshua B Tenenbaum · 2023
Later among the works it cites.
Efficient recovery learning using model predictive meta-reasoning
Shivam Vats, Maxim Likhachev, and Oliver Kroemer · 2023
Later among the works it cites.
Lotus: Continual imitation learning for robot manipulation through unsupervised skill discovery
Weikang Wan, Yifeng Zhu, Rutav Shah, and Yuke Zhu · 2023
Later among the works it cites.
Compositional Diffusion-Based Continuous Constraint Solvers
Zhutian Yang, Jiayuan Mao, Yilun Du, Jiajun Wu, Joshua B. Tenenbaum, Tomás Lozano-Pérez, and Leslie Pack Kaelbling · 2023
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
Tony Z Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn · 2023
Later among the works it cites.
Mobile aloha: Learning bimanual mobile manipulation with low-cost whole-body teleoperation
Zipeng Fu, Tony Z. Zhao, and Chelsea Finn · 2024
Closest in time.
Adaptive mobile manipulation for articulated objects in the open world
Haoyu Xiong, Russell Mendonca, Kenneth Shaw, and Deepak Pathak · 2024
Closest in time.