Fetching the paper…
Reading the bibliography…
A fundamental challenge in teaching robots is to provide an effective interface for human teachers to demonstrate useful skills to a robot.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in
1989
Earlier work this paper cites.
S. Schaal and C. Atkeson, “Robot juggling: implementation of memory-based learning,”
1994
Earlier work this paper cites.
N. Jakobi, P. Husbands, and I. Harvey, “Noise and the reality gap: The use of simulation in evolutionary robotics,” in
1995
Earlier work this paper cites.
L. P. Kaelbling, M. L. Littman, and A. W. Moore, “Reinforcement learning: A survey,”
1996
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in
2004
Earlier work this paper cites.
U. Syed and R. E. Schapire, “A game-theoretic approach to apprenticeship learning,” in
2007
Earlier work this paper cites.
I. Mordatch, Z. Popović, and E. Todorov, “Contact-invariant optimization for hand manipulation,” in
2012
Earlier work this paper cites.
V. Kumar, Y. Tassa, T. Erez, and E. Todorov, “Real-time behaviour synthesis for dynamic hand-manipulation,” in
2014
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,”
2015
Earlier work this paper cites.
V. Kumar and E. Todorov, “Mujoco haptix: A virtual reality system for hand manipulation,” in
2015
Earlier work this paper cites.
T. Zhang, G. Kahn, S. Levine, and P. Abbeel, “Learning deep control policies for autonomous aerial vehicles with mpc-guided policy search,” in
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot,
2016
Earlier work this paper cites.
L. Pinto and A. Gupta, “Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours,”
2016
Earlier work this paper cites.
S. Levine, P. Pastor, A. Krizhevsky, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,”
2016
Earlier work this paper cites.
F. Sadeghi and S. Levine, “Cad2rl: Real single-image flight without a single real image,”
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
R. Deimel and O. Brock, “A novel type of compliant and underactuated robotic hand for dexterous grasping,”
2016
Earlier work this paper cites.
V. Kumar, A. Gupta, E. Todorov, and S. Levine, “Learning dexterous manipulation policies from experience and imitation,”
2016
Earlier work this paper cites.
D. Gandhi, L. Pinto, and A. Gupta, “Learning to fly by crashing,” in
2017
Earlier work this paper cites.
J. Hwangbo, I. Sa, R. Siegwart, and M. Hutter, “Control of a quadrotor with reinforcement learning,”
2017
Earlier work this paper cites.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain randomization for transferring deep neural networks from simulation to the real world,” in
2017
Earlier work this paper cites.
J. Guo, S. Lu, H. Cai, W. Zhang, Y. Yu, and J. Wang, “Long text generation via adversarial training with leaked information,” in
2018
Earlier work this paper cites.
S. Gidaris and N. Komodakis, “Dynamic few-shot visual learning without forgetting,” in
2018
Earlier work this paper cites.
OpenAI, M. Andrychowicz, B. Baker, M. Chociej, R. Józefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba, “Learning dexterous in-hand manipulation,”
2018
Cited alongside, same era.
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine, “Learning complex dexterous manipulation with deep reinforcement learning and demonstrations,”
2018
Cited alongside, same era.
L. Pinto, M. Andrychowicz, P. Welinder, W. Zaremba, and P. Abbeel, “Asymmetric actor critic for image-based robot learning,”
2018
Cited alongside, same era.
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel, “Deep imitation learning for complex manipulation tasks from virtual reality teleoperation,” in
2018
Cited alongside, same era.
F. Zhang, V. Bazarevsky, A. Vakunov, A. Tkachenka, G. Sung, C.-L. Chang, and M. Grundmann, “Mediapipe hands: On-device real-time hand tracking,” 2020
2020
Later among the works it cites.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,”
2020
Later among the works it cites.
M. Huang, F. Li, W. Zou, and W. Zhang, “Sarg: A novel semi autoregressive generator for multi-turn incomplete utterance restoration,” in
2021
Later among the works it cites.
M. Caeiro-Rodríguez, I. Otero-González, F. A. Mikic-Fonte, and M. Llamas-Nistal, “A systematic review of commercial smart gloves: Current status and applications,”
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine, “Learning complex dexterous manipulation with deep reinforcement learning and demonstrations,” in
2018
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
M. Vecerik, O. Sushkov, D. Barker, T. Rothörl, T. Hester, and J. Scholz, “A practical approach to insertion with variable socket position using deep reinforcement learning,” in
2019
Cited alongside, same era.
OpenAI, I. Akkaya, M. Andrychowicz, M. Chociej, M. Litwin, B. McGrew, A. Petron, A. Paino, M. Plappert, G. Powell, R. Ribas, J. Schneider, N. Tezak, J. Tworek, P. Welinder, L. Weng, Q. Yuan, W. Zaremba, and L. Zhang, “Solving rubik’s cube with a robot hand,”
2019
Cited alongside, same era.
Y. Wu, W. Yan, T. Kurutach, L. Pinto, and P. Abbeel, “Learning to manipulate deformable objects without demonstrations,”
2019
Cited alongside, same era.
S. Li, X. Ma, H. Liang, M. Görner, P. Ruppel, B. Fang, F. Sun, and J. Zhang, “Vision-based teleoperation of shadow dexterous hand using end-to-end deep neural network,” in
2019
Cited alongside, same era.
Z. Gharaybeh, H. Chizeck, and A. Stewart, “Telerobotic control in virtual reality,” in
2019
Cited alongside, same era.
2021
Later among the works it cites.
P. Florence, C. Lynch, A. Zeng, O. Ramirez, A. Wahid, L. Downs, A. Wong, J. Lee, I. Mordatch, and J. Tompson, “Implicit behavioral cloning,” 2021
2021
Later among the works it cites.
J. Pari, N. M. Shafiullah, S. P. Arunachalam, and L. Pinto, “The surprising effectiveness of representation learning for visual imitation,” 2021
2021
Later among the works it cites.
I. Radosavovic, X. Wang, L. Pinto, and J. Malik, “State-only imitation learning for dexterous manipulation,” in
2021
Later among the works it cites.
X. Chen, S. Xie, and K. He, “An empirical study of training self-supervised vision transformers,” in
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Young, J. Pari, P. Abbeel, and L. Pinto, “Playful interactions for representation learning,”
2021
Later among the works it cites.
2022
Closest in time.
S. Gangapurwala, M. Geisert, R. Orsolino, M. Fallon, and I. Havoutis, “Rloc: Terrain-aware legged locomotion using reinforcement learning and optimal control,”
2022
Closest in time.
Y. Ma, F. Farshidian, T. Miki, J. Lee, and M. Hutter, “Combining learning-based locomotion policy with model-based manipulation for legged mobile manipulators,”
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
T. Hentschel and J. A. Neuhöfer, “Steady hands - an evaluation on the use of hand tracking in virtual reality training in nursing,” in
2022
Closest in time.
M. Salvato, N. Heravi, A. M. Okamura, and J. Bohg, “Predicting hand-object interaction for improved haptic feedback in mixed reality,”
2022
Closest in time.
2022
Closest in time.
L. Ericsson, H. Gouk, C. C. Loy, and T. M. Hospedales, “Self-supervised representation learning: Introduction, advances, and challenges,”
2022
Closest in time.