Fetching the paper…
Reading the bibliography…
Dexterous manipulation tasks involving contact-rich interactions pose a significant challenge for both model-based control systems and imitation learning algorithms.
Maximum margin planning
N. D. Ratliff, J. A. Bagnell, and M. Zinkevich · 2006
Earlier work this paper cites.
Policy search for motor primitives in robotics
J. Kober and J. Peters · 2008
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Contact-invariant optimization for hand manipulation
I. Mordatch, Z. Popović, and E. Todorov · 2012
Earlier work this paper cites.
Trajectories and keyframes for kinesthetic teaching: A human-robot interaction perspective
B. Akgun, M. Cakmak, J. W. Yoo, and A. L. Thomaz · 2012
Earlier work this paper cites.
A low-cost and modular, 20-dof anthropomorphic robotic hand: Design, actuation and modeling
Z. Xu, V. Kumar, and E. Todorov · 2013
Earlier work this paper cites.
Learning monocular reactive UAV control in cluttered natural environments
S. Ross, N. Melik-Barkhudarov, K. S. Shankar, A. Wendel, D. Dey, J. A. Bagnell, and M. Hebert · 2013
Earlier work this paper cites.
Real-time behaviour synthesis for dynamic hand-manipulation
V. Kumar, Y. Tassa, T. Erez, and E. Todorov · 2014
Earlier work this paper cites.
A direct method for trajectory optimization of rigid bodies through contact
M. Posa, C. Cantu, and R. Tedrake · 2014
Earlier work this paper cites.
Learning robot in-hand manipulation with tactile features
H. Van Hoof, T. Hermans, G. Neumann, and J. Peters · 2015
Earlier work this paper cites.
Maximum entropy deep inverse reinforcement learning
M. Wulfmeier, P. Ondruska, and I. Posner · 2015
Earlier work this paper cites.
Learning compound multi-step controllers under unknown dynamics
W. Han, S. Levine, and P. Abbeel · 2015
Earlier work this paper cites.
Learning dexterous manipulation policies from experience and imitation
V. Kumar, A. Gupta, E. Todorov, and S. Levine · 2016
Earlier work this paper cites.
A novel type of compliant and underactuated robotic hand for dexterous grasping
R. Deimel and O. Brock · 2016
Earlier work this paper cites.
Learning dexterous manipulation for a soft robotic hand from human demonstrations
A. Gupta, C. Eppner, S. Levine, and P. Abbeel · 2016
Earlier work this paper cites.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Earlier work this paper cites.
Imitation learning: A survey of learning methods
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne · 2017
Earlier work this paper cites.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine · 2017
Earlier work this paper cites.
Leave no trace: Learning to reset for safe and autonomous reinforcement learning
B. Eysenbach, S. Gu, J. Ibarz, and S. Levine · 2017
Earlier work this paper cites.
Improved training of wasserstein gans, 2017
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A. Courville · 2017
Cited alongside, same era.
Sim-to-real transfer of robotic control with dynamics randomization
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel · 2018
Cited alongside, same era.
Learning dexterous in-hand manipulation
OpenAI, M. Andrychowicz, B. Baker, M. Chociej, R. Józefowicz, B. McGrew, J. W. Pachocki, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba · 2018
Cited alongside, same era.
An algorithmic perspective on imitation learning
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, and J. Peters · 2018
Cited alongside, same era.
Reinforcement learning for non-prehensile manipulation: Transfer from simulation to physical system
K. Lowrey, S. Kolev, J. Dao, A. Rajeswaran, and E. Todorov · 2018
Learning precise 3d manipulation from multiple uncalibrated cameras
I. Akinola, J. Varley, and D. Kalashnikov · 2020
Later among the works it cites.
Better-than-demonstrator imitation learning via automatically-ranked demonstrations
D. S. Brown, W. Goo, and S. Niekum · 2020
Later among the works it cites.
Image augmentation is all you need: Regularizing deep reinforcement learning from pixels
I. Kostrikov, D. Yarats, and R. Fergus · 2020
Later among the works it cites.
Offline learning from demonstrations and unlabeled experience
K. Zolna, A. Novikov, K. Konyushkova, C. Gulcehre, Z. Wang, Y. Aytar, M. Denil, N. de Freitas, and S. Reed · 2020
Later among the works it cites.
Continual learning of control primitives: Skill discovery via reset-games
K. Xu, S. Verma, C. Finn, and S. Levine · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Including uncertainty when learning from human corrections
D. P. Losey and M. K. O’Malley · 2018
Cited alongside, same era.
Active reward learning from critiques
Y. Cui and S. Niekum · 2018
Cited alongside, same era.
Guiding policies with language via meta-learning
J. D. Co-Reyes, A. Gupta, S. Sanjeev, N. Altieri, J. DeNero, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Variational inverse control with events: A general framework for data-driven reward definition
J. Fu, A. Singh, D. Ghosh, L. Yang, and S. Levine · 2018
Cited alongside, same era.
Survey on human–robot collaboration in industrial settings: Safety, intuitive interfaces and applications
V. Villani, F. Pini, F. Leali, and C. Secchi · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Cited alongside, same era.
Soft actor-critic algorithms and applications
T. Haarnoja, A. Zhou, K. Hartikainen, G. Tucker, S. Ha, J. Tan, V. Kumar, H. Zhu, A. Gupta, P. Abbeel, et al · 2018
Cited alongside, same era.
Later among the works it cites.
Human-in-the-loop imitation learning using remote teleoperation, 2020
A. Mandlekar, D. Xu, R. Martín-Martín, Y. Zhu, L. Fei-Fei, and S. Savarese · 2020
Later among the works it cites.
Learning compliant grasping and manipulation by teleoperation with adaptive force control, 2021
C. Zeng, S. Li, Y. Jiang, Q. Li, Z. Chen, C. Yang, and J. Zhang · 2021
Later among the works it cites.
Transferring dexterous manipulation from gpu simulation to a remote real-world trifinger
A. Allshire, M. Mittal, V. Lodaya, V. Makoviychuk, D. Makoviichuk, F. Widmaier, M. Wüthrich, S. Bauer, A. Handa, and A. Garg · 2021
Later among the works it cites.
A. Gupta, J. Yu, T. Z. Zhao, V. Kumar, A. Rovinsky, K. Xu, T. Devlin, and S. Levine · 2021
Later among the works it cites.
Autonomous reinforcement learning via subgoal curricula, 2021
A. Sharma, A. Gupta, S. Levine, K. Hausman, and C. Finn · 2021
Later among the works it cites.
Multifingered robot hand compliant manipulation based on vision-based demonstration and adaptive force control
C. Zeng, S. Li, Z. Chen, C. Yang, F. Sun, and J. Zhang · 2022
Later among the works it cites.
Holo-dex: Teaching dexterity with immersive mixed reality, 2022
S. P. Arunachalam, I. Güzey, S. Chintala, and L. Pinto · 2022
Later among the works it cites.
Dexterous manipulation from images: Autonomous real-world rl via substep guidance, 2022
K. Xu, Z. Hu, R. Doshi, A. Rovinsky, V. Kumar, A. Gupta, and S. Levine · 2022
Later among the works it cites.
Learning multimodal rewards from rankings
V. Myers, E. Biyik, N. Anari, and D. Sadigh · 2022
Later among the works it cites.
Autonomous reinforcement learning: Formalism and benchmarking, 2022
A. Sharma, K. Xu, N. Sardana, A. Gupta, K. Hausman, S. Levine, and C. Finn · 2022
Later among the works it cites.
Hybrid rl: Using both offline and online data can make rl efficient
Y. Song, Y. Zhou, A. Sekhari, J. A. Bagnell, A. Krishnamurthy, and W. Sun · 2022
Later among the works it cites.
Efficient online reinforcement learning with offline data, 2023
P. J. Ball, L. Smith, I. Kostrikov, and S. Levine · 2023
Closest in time.
All the feels: A dexterous hand with large area sensing, 2023
R. Bhirangi, A. DeFranco, J. Adkins, C. Majidi, A. Gupta, T. Hellebrekers, and V. Kumar · 2023
Closest in time.
Learning and adapting agile locomotion skills by transferring experience, 2023
L. Smith, J. C. Kew, T. Li, L. Luu, X. B. Peng, S. Ha, J. Tan, and S. Levine · 2023
Closest in time.
Self-improving robots: End-to-end autonomous visuomotor reinforcement learning, 2023
A. Sharma, A. M. Ahmed, R. Ahmad, and C. Finn · 2023
Closest in time.