Fetching the paper…
Reading the bibliography…
We need to look at our shoelaces as we first learn to tie them but having mastered this skill, can do it from touch alone.
The role of tutoring in problem solving
David Wood, Jerome S Bruner, and Gail Ross · 1976
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L Littman, and Anthony R Cassandra · 1998
Earlier work this paper cites.
Transfer learning via inter-task mappings for temporal difference learning
Matthew E Taylor, Peter Stone, and Yaxin Liu · 2007
Earlier work this paper cites.
A new learning paradigm: Learning using privileged information
Vladimir Vapnik and Akshay Vashist · 2009
Earlier work this paper cites.
Interaction between learning and development
Lev Vygotsky et al · 2011
Earlier work this paper cites.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine · 2017
Earlier work this paper cites.
Visual closed-loop control for pouring liquids
Connor Schenck and Dieter Fox · 2017
Earlier work this paper cites.
Domain randomization for transferring deep neural networks from simulation to the real world
Josh Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel · 2017
Earlier work this paper cites.
Learning dexterous in-hand manipulation
Marcin Andrychowicz, Bowen Baker, Maciek Chociej, Rafal Jozefowicz, Bob McGrew, Jakub Pachocki, Arthur Petron, Matthias Plappert, Glenn Powell, Alex Ray, et al · 2018
Earlier work this paper cites.
Closing the sim-to-real loop: Adapting simulation randomization with real world experience
Yevgen Chebotar, Ankur Handa, Viktor Makoviychuk, Miles Macklin, Jan Issac, Nathan Ratliff, and Dieter Fox · 2018
Earlier work this paper cites.
Variational inverse control with events: A general framework for data-driven reward definition
Justin Fu, Avi Singh, Dibya Ghosh, Larry Yang, and Sergey Levine · 2018
Earlier work this paper cites.
Deepmimic: Example-guided deep reinforcement learning of physics-based character skills
Xue Bin Peng, Pieter Abbeel, Sergey Levine, and Michiel Van de Panne · 2018
Earlier work this paper cites.
Asymmetric actor critic for image-based robot learning
Lerrel Pinto, Marcin Andrychowicz, Peter Welinder, Wojciech Zaremba, and Pieter Abbeel · 2018
Earlier work this paper cites.
Time-contrastive networks: Self-supervised learning from video
Pierre Sermanet, Corey Lynch, Yevgen Chebotar, Jasmine Hsu, Eric Jang, Stefan Schaal, and Sergey Levine · 2018
Cited alongside, same era.
Simultaneously learning vision and feature-based control policies for real-world ball-in-a-cup
Devin Schwab, Tobias Springenberg, Murilo F Martins, Thomas Lampe, Michael Neunert, Abbas Abdolmaleki, Tim Hertweck, Roland Hafner, Francesco Nori, and Martin Riedmiller · 2019
Cited alongside, same era.
Learning by cheating
Dian Chen, Brady Zhou, Vladlen Koltun, and Philipp Krähenbühl · 2020
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
Danijar Hafner, Timothy Lillicrap, Jimmy Ba, and Mohammad Norouzi · 2020
Cited alongside, same era.
Privileged information dropout in reinforcement learning
Pierre-Alexandre Kamienny, Kai Arulkumaran, Feryal Behbahani, Wendelin Boehmer, and Shimon Whiteson · 2020
Cited alongside, same era.
Know thyself: Transferable visual control policies through robot-awareness
Edward S. Hu, Kun Huang, Oleh Rybkin, and Dinesh Jayaraman · 2022
Later among the works it cites.
Leveraging fully observable policies for learning under partial observability
Hai Nguyen, Andrea Baisero, Dian Wang, Christopher Amato, and Robert Platt · 2022
Later among the works it cites.
Transfer RL across observation feature spaces via model-based regularization
Yanchao Sun, Ruijie Zheng, Xiyao Wang, Andrew E Cohen, and Furong Huang · 2022
Later among the works it cites.
Sequence model imitation learning with unobserved contexts
Gokul Swamy, Sanjiban Choudhury, J Bagnell, and Steven Z Wu · 2022
Later among the works it cites.
Impossibly good experts and how to follow them
Aaron Walsman, Muru Zhang, Sanjiban Choudhury, Dieter Fox, and Ali Farhadi · 2022
Later among the works it cites.
Sequential dexterity: Chaining dexterous policies for long-horizon manipulation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning Quadrupedal Locomotion over Challenging Terrain
Joonho Lee, Jemin Hwangbo, Lorenz Wellhausen, Vladlen Koltun, and Marco Hutter · 2020
Cited alongside, same era.
Unbiased asymmetric reinforcement learning under partial observability
Andrea Baisero and Christopher Amato · 2021
Cited alongside, same era.
RMA: Rapid Motor Adaptation for Legged Robots
Ashish Kumar, Zipeng Fu, Deepak Pathak, and Jitendra Malik · 2021
Cited alongside, same era.
Attention-privileged reinforcement learning
Sasha Salter, Dushyant Rao, Markus Wulfmeier, Raia Hadsell, and Ingmar Posner · 2021
Cited alongside, same era.
Reinforcement learning using guided observability
Stephan Weigand, Pascal Klink, Jan Peters, and Joni Pajarinen · 2021
Cited alongside, same era.
Bridging the imitation gap by adaptive insubordination
Luca Weihs, Unnat Jain, Iou-Jen Liu, Jordi Salvador, Svetlana Lazebnik, Aniruddha Kembhavi, and Alexander Schwing · 2021
Cited alongside, same era.
Policy transfer across visual and dynamics domain gaps via iterative grounding
Grace Zhang, Linghan Zhong, Youngwoon Lee, and Joseph J Lim · 2021
Cited alongside, same era.
Yuanpei Chen, Chen Wang, Li Fei-Fei, and C Karen Liu · 2023
Later among the works it cites.
Mastering diverse domains through world models
Danijar Hafner, Jurgis Pasukonis, Jimmy Ba, and Timothy Lillicrap · 2023
Later among the works it cites.
Watch and match: Supercharging imitation with regularized optimal transport
Siddhant Haldar, Vaibhav Mathur, Denis Yarats, and Lerrel Pinto · 2023
Later among the works it cites.
Informed POMDP: Leveraging Additional Information in Model-Based RL
Gaspard Lambrechts, Adrien Bolland, and Damien Ernst · 2023
Later among the works it cites.
In-hand object rotation via rapid motor adaptation
Haozhi Qi, Ashish Kumar, Roberto Calandra, Yi Ma, and Jitendra Malik · 2023
Later among the works it cites.
Multi-view masked world models for visual robotic manipulation
Younggyo Seo, Junsu Kim, Stephen James, Kimin Lee, Jinwoo Shin, and Pieter Abbeel · 2023
Later among the works it cites.
Tgrl: An algorithm for teacher guided reinforcement learning
Idan Shenfeld, Zhang-Wei Hong, Aviv Tamar, and Pulkit Agrawal · 2023
Later among the works it cites.
Transferring implicit knowledge of non-visual object properties across heterogeneous robot morphologies
Gyan Tatiya, Jonathan Francis, and Jivko Sinapov · 2023
Later among the works it cites.