Fetching the paper…
Reading the bibliography…
We consider how to most efficiently leverage teleoperator time to collect data for learning robust image-based value functions and policies for sparse reward robotic tasks.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in NIPS
1988
Earlier work this paper cites.
S. Schaal, “Learning from demonstration,” Advances in neural information processing systems
1996
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,” Robotics and autonomous systems
2009
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics
2011
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” arXiv preprint arXiv:1412.6980
2014
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition
2016
Earlier work this paper cites.
J. Bradbury, R. Frostig, P. Hawkins, M. J. Johnson, C. Leary, D. Maclaurin, G. Necula, A. Paszke, J. VanderPlas, S. Wanderman-Milne, and Q. Zhang, “JAX: composable transformations of Python+NumPy programs,” 2018
2018
Earlier work this paper cites.
S. Fujimoto, D. Meger, and D. Precup, “Off-policy deep reinforcement learning without exploration,” in International conference on machine learning
2019
Earlier work this paper cites.
M. Kelly, C. Sidrane, K. Driggs-Campbell, and M. J. Kochenderfer, “Hg-dagger: Interactive imitation learning with human experts,” in 2019 International Conference on Robotics and Automation (ICRA)
2019
Cited alongside, same era.
J. Chen and N. Jiang, “Information-theoretic considerations in batch reinforcement learning,” in International Conference on Machine Learning
2019
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. Kumar, J. Hong, A. Singh, and S. Levine, “Should i run offline reinforcement learning or behavioral cloning?,” in Deep RL Workshop NeurIPS 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
R. Hoque, A. Balakrishna, C. Putterman, M. Luo, D. S. Brown, D. Seita, B. Thananjeyan, E. Novoseller, and K. Goldberg, “Lazydagger: Reducing context switching in interactive imitation learning,” in 2021 IEEE 17th International Conference on Automation Science and Engineering (CASE)
2021
Later among the works it cites.
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet, “Learning latent plans from play,” in Conference on robot learning
2020
Cited alongside, same era.
J. Heek, A. Levskaya, A. Oliver, M. Ritter, B. Rondepierre, A. Steiner, and M. van Zee, “Flax: A neural network library and ecosystem for JAX,” 2020
2020
Cited alongside, same era.
2021
Cited alongside, same era.
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn, “Bc-z: Zero-shot task generalization with robotic imitation learning,” in Conference on Robot Learning
2022
Closest in time.
2022
Closest in time.
A. Wong, A. Zeng, A. Bose, A. Wahid, D. Kalashnikov, I. Krasin, J. Varley, J. Lee, J. Tompson, M. Attarian, P. Florence, R. Baruch, S. Xu, S. Welker, V. Sindhwani, V. Vanhoucke, and W. Gramlich, “Pyreach - python client sdk for robot remote control.” https://github.com/google-research/pyreach , 2022
2022
Closest in time.