Fetching the paper…
Reading the bibliography…
A common strategy in modern learning systems is to learn a representation that is useful for many tasks, a.k.a.
Efficient training of artificial neural networks for autonomous navigation
D. A. Pomerleau · 1991
Earlier work this paper cites.
A model of inductive bias learning
Jonathan Baxter · 2000
Earlier work this paper cites.
Rademacher and gaussian complexities: Risk bounds and structural results
Peter L. Bartlett and Shahar Mendelson · 2003
Earlier work this paper cites.
On the sample complexity of reinforcement learning
Sham Machandranath Kakade et al · 2003
Earlier work this paper cites.
Error bounds for approximate value iteration
Rémi Munos · 2005
Earlier work this paper cites.
Finite-time bounds for fitted value iteration
Rémi Munos and Csaba Szepesvári · 2008
Earlier work this paper cites.
Search-based structured prediction
Hal Daumé, Iii, John Langford, and Daniel Marcu · 2009
Earlier work this paper cites.
Transfer bounds for linear feature learning
Andreas Maurer · 2009
Earlier work this paper cites.
Efficient reductions for imitation learning
Stéphane Ross and Drew Bagnell · 2010
Earlier work this paper cites.
A reduction from apprenticeship learning to classification
Umar Syed and Robert E Schapire · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Y. Bengio, Aaron Courville, and Pascal Vincent · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Reinforcement and imitation learning via interactive no-regret learning
Stéphane Ross and J. Andrew Bagnell · 2014
Cited alongside, same era.
Learning to search better than your teacher
Kai-Wei Chang, Akshay Krishnamurthy, Alekh Agarwal, Hal Daumé, III, and John Langford · 2015
Cited alongside, same era.
Openai gym, 2016
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
The benefit of multitask representation learning
Task-embedded control networks for few-shot imitation learning
Stephen James, Michael Bloesch, and Andrew Davison · 2018
Later among the works it cites.
Agile autonomous driving using end-to-end deep imitation learning
Yunpeng Pan, Ching-An Cheng, Kamil Saigol, Keuntaek Lee, Xinyan Yan, Evangelos Theodorou, and Byron Boots · 2018
Later among the works it cites.
Truncated horizon policy search: Combining reinforcement learning and imitation learning
Wen Sun, J. Andrew Bagnell, and Byron Boots · 2018
Later among the works it cites.
Behavioral cloning from observation
Faraz Torabi, Garrett Warnell, and Peter Stone · 2018
Later among the works it cites.
A theoretical analysis of contrastive unsupervised representation learning
Sanjeev Arora, Hrishikesh Khandeparkar, Mikhail Khodak, Orestis Plevrakis, and Nikunj Saunshi · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Andreas Maurer, Massimiliano Pontil, and Bernardino Romera-Paredes · 2016
Cited alongside, same era.
Openai baselines
Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, Yuhuai Wu, and Peter Zhokhov · 2017
Cited alongside, same era.
One-shot imitation learning
Yan Duan, Marcin Andrychowicz, Bradly Stadie, OpenAI Jonathan Ho, Jonas Schneider, Ilya Sutskever, Pieter Abbeel, and Wojciech Zaremba · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Complete dictionary recovery over the sphere I: Overview and the geometric picture
Ju Sun, Qing Qu, and John Wright · 2017
Cited alongside, same era.
Learning to see physics via visual de-animation
Jiajun Wu, Erika Lu, Pushmeet Kohli, Bill Freeman, and Joshua B. Tenenbaum · 2017
Cited alongside, same era.
Incremental learning-to-learn with statistical guarantees
Giulia Denevi, Carlo Ciliberto, Dimitris Stamos, and Massimiliano Pontil · 2018
Cited alongside, same era.
Generalize across tasks: Efficient algorithms for linear representation learning
Brian Bullins, Elad Hazan, Adam Kalai, and Roi Livni · 2019
Later among the works it cites.
Learning-to-learn stochastic gradient descent with biased regularization
Giulia Denevi, Carlo Ciliberto, Riccardo Grazzi, and Massimiliano Pontil · 2019
Later among the works it cites.
Online meta-learning
Chelsea Finn, Aravind Rajeswaran, Sham Kakade, and Sergey Levine · 2019
Later among the works it cites.
Adaptive gradient-based meta-learning methods
Mikhail Khodak, Maria-Florina Balcan, and Ameet Talwalkar · 2019
Later among the works it cites.
Batch policy learning under constraints
Hoang Le, Cameron Voloshin, and Yisong Yue · 2019
Later among the works it cites.
Rapid learning or feature reuse? towards understanding the effectiveness of maml
Aniruddh Raghu, Maithra Raghu, Samy Bengio, and Oriol Vinyals · 2019
Later among the works it cites.
Meta-learning with implicit gradients
Aravind Rajeswaran, Chelsea Finn, Sham Kakade, and Sergey Levine · 2019
Later among the works it cites.
Provably efficient imitation learning from observation alone
Wen Sun, Anirudh Vemula, Byron Boots, and J Andrew Bagnell · 2019
Later among the works it cites.