Fetching the paper…
Reading the bibliography…
Imitation learning considerably simplifies policy synthesis compared to alternative approaches by exploiting access to expert demonstrations.
Efficient training of artificial neural networks for autonomous navigation
Dean A Pomerleau · 1991
Earlier work this paper cites.
’neural-gas’ network for vector quantization and its application to time-series prediction
T.M. Martinetz, S.G. Berkovich, and K.J. Schulten · 1993
Earlier work this paper cites.
A growing neural gas network learns topologies
Bernd Fritzke · 1994
Earlier work this paper cites.
An incremental growing neural gas learns topologies
Y. Prudent and A. Ennaji · 2005
Earlier work this paper cites.
Efficient reductions for imitation learning
Stéphane Ross and Drew Bagnell · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Model-based reinforcement learning and the eluder dimension
Ian Osband and Benjamin Van Roy · 2014
Earlier work this paper cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Earlier work this paper cites.
Robust reinforcement learning with relevance vector machines
Minwoo Lee and Charles W Anderson · 2016
Earlier work this paper cites.
Carla: An open urban driving simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun · 2017
Earlier work this paper cites.
Dart: Noise injection for robust imitation learning
Michael Laskey, Jonathan Lee, Roy Fox, Anca Dragan, and Ken Goldberg · 2017
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Q-learning with nearest neighbors
Devavrat Shah and Qiaomin Xie · 2018
Earlier work this paper cites.
Exploring the limitations of behavior cloning for autonomous driving
Felipe Codevilla, Eder Santana, Antonio M López, and Adrien Gaidon · 2019
Earlier work this paper cites.
Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning
Abhishek Gupta, Vikash Kumar, Corey Lynch, Sergey Levine, and Karol Hausman · 2019
Cited alongside, same era.
Advantage-weighted regression: Simple and scalable off-policy reinforcement learning
Xue Bin Peng, Aviral Kumar, Grace Zhang, and Sergey Levine · 2019
Cited alongside, same era.
D4rl: Datasets for deep data-driven reinforcement learning
Justin Fu, Aviral Kumar, Ofir Nachum, George Tucker, and Sergey Levine · 2020
Cited alongside, same era.
Conservative q-learning for offline reinforcement learning
Aviral Kumar, Aurick Zhou, George Tucker, and Sergey Levine · 2020
Cited alongside, same era.
Toward the fundamental limits of imitation learning
Nived Rajaraman, Lin F. Yang, Jiantao Jiao, and Kannan Ramchandran · 2020
Transformers are sample efficient world models
Vincent Micheli, Eloi Alonso, and François Fleuret · 2022
Later among the works it cites.
Behavior transformers: Cloning k k modes with one stone
Nur Muhammad Shafiullah, Zichen Cui, Ariuntuya Arty Altanzaya, and Lerrel Pinto · 2022
Later among the works it cites.
Improving neural network robustness via persistency of excitation
Kaustubh Sridhar, Oleg Sokolsky, Insup Lee, and James Weimer · 2022
Later among the works it cites.
Diffusion policies as an expressive policy class for offline reinforcement learning
Zhendong Wang, Jonathan J Hunt, and Mingyuan Zhou · 2022
Later among the works it cites.
Interpretable detection of distribution shifts in learning enabled cyber-physical systems
Yahan Yang, Ramneet Kaur, Souradeep Dutta, and Insup Lee · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning invariant representations for reinforcement learning without reconstruction
Amy Zhang, Rowan McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine · 2020
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
Patrick Esser, Robin Rombach, and Bjorn Ommer · 2021
Cited alongside, same era.
A minimalist approach to offline reinforcement learning
Scott Fujimoto and Shixiang Shane Gu · 2021
Cited alongside, same era.
What matters in learning from offline human demonstrations for robot manipulation
Ajay Mandlekar, Danfei Xu, Josiah Wong, Soroush Nasiriany, Chen Wang, Rohun Kulkarni, Li Fei-Fei, Silvio Savarese, Yuke Zhu, and Roberto Martín-Martín · 2021
Cited alongside, same era.
The surprising effectiveness of representation learning for visual imitation
Jyothish Pari, Nur Muhammad Shafiullah, Sridhar Pandian Arunachalam, and Lerrel Pinto · 2021
Cited alongside, same era.
On the value of interaction and function approximation in imitation learning
Nived Rajaraman, Yanjun Han, Lin Yang, Jingbo Liu, Jiantao Jiao, and Kannan Ramchandran · 2021
Cited alongside, same era.
Memory classifiers: Two-stage classification for robustness in machine learning
Souradeep Dutta, Yahan Yang, Elena Bernardis, Edgar Dobriban, and Insup Lee · 2022
Cited alongside, same era.
Diffusion policy: Visuomotor policy learning via action diffusion
Cheng Chi, Siyuan Feng, Yilun Du, Zhenjia Xu, Eric Cousineau, Benjamin Burchfiel, and Shuran Song · 2023
Closest in time.
Exploring with sticky mittens: Reinforcement learning with expert interventions via option templates
Souradeep Dutta, Kaustubh Sridhar, Osbert Bastani, Edgar Dobriban, James Weimer, Insup Lee, and Julia Parish-Morris · 2023
Closest in time.
Memory classifiers for robust ecg classification against physiological noise
Kuk Jin Jang, Souradeep Dutta, Jean Park, James Weimer, and Insup Lee · 2023
Closest in time.
Incremental anomaly detection with guarantee in the internet of medical things
Xiayan Ji, Hyonyoung Choi, Oleg Sokolsky, and Insup Lee · 2023
Closest in time.
When do transformers shine in rl? decoupling memory from credit assignment
Tianwei Ni, Michel Ma, Benjamin Eysenbach, and Pierre-Luc Bacon · 2023
Closest in time.
Imitating human behaviour with diffusion models
Tim Pearce, Tabish Rashid, Anssi Kanervisto, Dave Bignell, Mingfei Sun, Raluca Georgescu, Sergio Valcarcel Macua, Shan Zheng Tan, Ida Momennejad, Katja Hofmann, et al · 2023
Closest in time.
Offlinerl-kit: An elegant pytorch offline reinforcement learning library
Yihao Sun · 2023
Closest in time.
Incremental learning with memory regressors for motion prediction in autonomous racing
Yahan Yang, Souradeep Dutta, Kuk Jin Jang, Oleg Sokolsky, and Insup Lee · 2023
Closest in time.
Learning fine-grained bimanual manipulation with low-cost hardware
Tony Z Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn · 2023
Closest in time.