Fetching the paper…
Reading the bibliography…
We tackle real-world long-horizon robot manipulation tasks through skill discovery.
“The senses considered as perceptual systems”
James Gibson and Leonard Carmichael · 1966
Earlier work this paper cites.
“A robust layered control system for a mobile robot”
Rodney Brooks · 1986
Earlier work this paper cites.
“A unified approach for motion and force control of robot manipulators: The operational space formulation”
Oussama Khatib · 1987
Earlier work this paper cites.
“Intelligence without representation”
Rodney Brooks · 1991
Earlier work this paper cites.
“Skill acquisition from human demonstration using a hidden markov model”
Geir Hovland, Pavan Sikka and Brenan McCarragher · 1996
Earlier work this paper cites.
“The MAXQ Method for Hierarchical Reinforcement Learning.”
Thomas Dietterich · 1998
Earlier work this paper cites.
“Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning”
Richard Sutton, Doina Precup and Satinder Singh · 1999
Earlier work this paper cites.
“Temporal Abstraction in Reinforcement Learning”, 2000
Doina Precup · 2000
Earlier work this paper cites.
“Event structure in perception and conception”
Jeffrey Zacks and Barbara Tversky · 2001
Earlier work this paper cites.
“Training products of experts by minimizing contrastive divergence”
Geoffrey Hinton · 2002
Earlier work this paper cites.
“A hierarchical architecture for behavior-based robots”
Monica Nicolescu and Maja Matarić · 2002
Earlier work this paper cites.
“A tutorial on spectral clustering”
Ulrike Von · 2007
Earlier work this paper cites.
“Skill discovery in continuous reinforcement learning domains using skill chaining”
George Konidaris and Andrew Barto · 2009
Earlier work this paper cites.
“Constructing Skill Trees for Reinforcement Learning Agents from Demonstration Trajectories.”
George Konidaris, Scott Kuindersma, Andrew Barto and Roderic Grupen · 2010
Earlier work this paper cites.
“Hierarchical task and motion planning in the now”
Leslie Kaelbling and Tomás Lozano-Pérez · 2011
Earlier work this paper cites.
“Learning and generalization of complex tasks from unstructured demonstrations”
Scott Niekum, Sarah Osentoski, George Konidaris and Andrew Barto · 2012
Earlier work this paper cites.
“Integrated task and motion planning in belief space”
Leslie Kaelbling and Tomás Lozano-Pérez · 2013
Earlier work this paper cites.
“Auto-encoding variational bayes”
Diederik Kingma and Max Welling · 2013
Earlier work this paper cites.
“Action recognition by hierarchical mid-level action elements”
Tian Lan, Yuke Zhu, Amir Zamir and Silvio Savarese · 2015
Earlier work this paper cites.
“Online bayesian changepoint detection for articulated motion models”
Scott Niekum, Sarah Osentoski, Christopher Atkeson and Andrew Barto · 2015
Cited alongside, same era.
“Universal value function approximators”
Tom Schaul, Daniel Horgan, Karol Gregor and David Silver · 2015
Cited alongside, same era.
“A bottom-up approach for pancreas segmentation using cascaded superpixels and (deep) image patch labeling”
Amal Farag et al · 2016
Cited alongside, same era.
“Variational intrinsic control”
Karol Gregor, Danilo Rezende and Daan Wierstra · 2016
Cited alongside, same era.
“Stacked hourglass networks for human pose estimation”
Alejandro Newell, Kaiyu Yang and Jia Deng · 2016
Cited alongside, same era.
“Hindsight experience replay”
Marcin Andrychowicz et al · 2017
“One-shot hierarchical imitation learning of compound visuomotor tasks”
Tianhe Yu, Pieter Abbeel, Sergey Levine and Chelsea Finn · 2018
Later among the works it cites.
“Deep imitation learning for complex manipulation tasks from virtual reality teleoperation”
Tianhao Zhang et al · 2018
Later among the works it cites.
“Option discovery using deep skill chaining”
Akhil Bagaria and George Konidaris · 2019
Later among the works it cites.
“Real-time Multisensory Affordance-based Control for Adaptive Object Manipulation”
Vivian Chu et al · 2019
Later among the works it cites.
“Efficient parameter-free clustering using first neighbor relations”
Saquib Sarfraz, Vivek Sharma and Rainer Stiefelhagen · 2019
Later among the works it cites.
“Dynamics-Aware Unsupervised Discovery of Skills”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Yan Duan et al · 2017
Cited alongside, same era.
“Multi-level discovery of deep options”
Roy Fox, Sanjay Krishnan, Ion Stoica and Ken Goldberg · 2017
Cited alongside, same era.
“Transition state clustering: Unsupervised surgical trajectory segmentation for robot learning”
Sanjay Krishnan et al · 2017
Cited alongside, same era.
“Feudal networks for hierarchical reinforcement learning”
Alexander Vezhnevets et al · 2017
Cited alongside, same era.
“Diversity is all you need: Learning skills without a reward function”
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz and Sergey Levine · 2018
Cited alongside, same era.
“Learning an embedding space for transferable robot skills”
Karol Hausman et al · 2018
Cited alongside, same era.
Archit Sharma et al · 2019
Later among the works it cites.
“Avid: Learning multi-stage tasks via pixel-level translation of human videos”
Laura Smith et al · 2019
Later among the works it cites.
“Bottom-up object detection by grouping extreme and center points”
Xingyi Zhou, Jiacheng Zhuo and Philipp Krahenbuhl · 2019
Later among the works it cites.
“Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning”
Abhishek Gupta et al · 2020
Later among the works it cites.
“Making sense of vision and touch: Learning multimodal representations for contact-rich tasks”
Michelle Lee et al · 2020
Later among the works it cites.
“Iris: Implicit reinforcement without interaction at scale for learning control from offline robot manipulation data”
Ajay Mandlekar et al · 2020
Later among the works it cites.
“Learning to generalize across long-horizon tasks from human demonstrations”
Ajay Mandlekar et al · 2020
Later among the works it cites.
“Modeling Long-horizon Tasks as Sequential Interaction Landscapes”
Sören Pirk, Karol Hausman, Alexander Toshev and Mohi Khansari · 2020
Later among the works it cites.
“Learning robot skills with temporal variational inference”
Tanmay Shankar and Abhinav Gupta · 2020
Later among the works it cites.
“robosuite: A modular simulation framework and benchmark for robot learning”
Yuke Zhu, Josiah Wong, Ajay Mandlekar and Roberto Martı́n-Martı́n · 2020
Later among the works it cites.
“Integrated task and motion planning”
Caelan Garrett et al · 2021
Closest in time.
“Temporally-Weighted Hierarchical Clustering for Unsupervised Action Segmentation”
M Sarfraz et al · 2021
Closest in time.
“SKID RAW: Skill Discovery from Raw Trajectories”
Daniel Tanneberg, Kai Ploeger, Elmar Rueckert and Jan Peters · 2021
Closest in time.
“Learning Multi-Arm Manipulation Through Collaborative Teleoperation”
Albert Tung et al · 2021
Closest in time.