Fetching the paper…
Reading the bibliography…
When autonomous vehicles are deployed on public roads, they will encounter countless and diverse driving situations.
A software package for sequential quadratic programming
Dieter Kraft · 1988
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Earlier work this paper cites.
The construction of movement with behavior-specific and behavior-independent modules
Jian Jing, Elizabeth C Cropper, Itay Hurwitz, and Klaudiusz R Weiss · 2004
Earlier work this paper cites.
Reinforcement learning of motor skills with policy gradients
Jan Peters and Stefan Schaal · 2008
Earlier work this paper cites.
Hierarchically organized behavior and its neural foundations: a reinforcement learning perspective
Matthew M Botvinick, Yael Niv, and Andew G Barto · 2009
Earlier work this paper cites.
Learning motor primitives for robotics
Jens Kober and Jan Peters · 2009
Earlier work this paper cites.
Accelerating reinforcement learning with learned skill priors
Karl Pertsch, Youngwoon Lee, and Joseph J Lim · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Robot learning from demonstration by constructing skill trees
George Konidaris, Scott Kuindersma, Roderic Grupen, and Andrew Barto · 2012
Earlier work this paper cites.
Learning to drive using inverse reinforcement learning and deep q-networks
Sahand Sharifzadeh, Ioannis Chiotellis, Rudolph Triebel, and Daniel Cremers · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Earlier work this paper cites.
The option-critic architecture
Pierre-Luc Bacon, Jean Harb, and Doina Precup · 2017
Earlier work this paper cites.
Improved trajectory planning for on-road self-driving vehicles via combined graph search, optimization & topology analysis
Tianyu Gu · 2017
Earlier work this paper cites.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards
Mel Vecerik, Todd Hester, Jonathan Scholz, Fumin Wang, Olivier Pietquin, Bilal Piot, Nicolas Heess, Thomas Rothörl, Thomas Lampe, and Martin Riedmiller · 2017
Cited alongside, same era.
How would surround vehicles move? a unified framework for maneuver classification and motion prediction
Nachiket Deo, Akshay Rangesh, and Mohan M Trivedi · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Deep q-learning from demonstrations
Todd Hester, Matej Vecerik, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Ian Osband, et al · 2018
Cited alongside, same era.
Compositional imitation learning: Explaining and executing one task at a time
Awac: Accelerating online reinforcement learning with offline datasets
Ashvin Nair, Abhishek Gupta, Murtaza Dalal, and Sergey Levine · 2020
Later among the works it cites.
Efficient uncertainty-aware decision-making for automated driving using guided branching
Lu Zhang, Wenchao Ding, Jing Chen, and Shaojie Shen · 2020
Later among the works it cites.
Interpretable end-to-end urban autonomous driving with latent deep reinforcement learning
Jianyu Chen, Shengbo Eben Li, and Masayoshi Tomizuka · 2021
Later among the works it cites.
Accelerating robotic reinforcement learning via parameterized action primitives
Murtaza Dalal, Deepak Pathak, and Russ R Salakhutdinov · 2021
Later among the works it cites.
Epsilon: An efficient planning system for automated vehicles in highly interactive environments
Wenchao Ding, Lu Zhang, Jing Chen, and Shaojie Shen · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Thomas Kipf, Yujia Li, Hanjun Dai, Vinicius Zambaldi, Edward Grefenstette, Pushmeet Kohli, and Peter Battaglia · 2018
Cited alongside, same era.
Cirl: Controllable imitative reinforcement learning for vision-based self-driving
Xiaodan Liang, Tairui Wang, Luona Yang, and Eric Xing · 2018
Cited alongside, same era.
Neural probabilistic motor primitives for humanoid control
Josh Merel, Leonard Hasenclever, Alexandre Galashov, Arun Ahuja, Vu Pham, Greg Wayne, Yee Whye Teh, and Nicolas Heess · 2018
Cited alongside, same era.
Overcoming exploration in reinforcement learning with demonstrations
Ashvin Nair, Bob McGrew, Marcin Andrychowicz, Wojciech Zaremba, and Pieter Abbeel · 2018
Cited alongside, same era.
The promise of hierarchical reinforcement learning
Yannis Flet-Berliac · 2019
Cited alongside, same era.
Variable impedance control in end-effector space: An action space for reinforcement learning in contact-rich tasks
Roberto Martín-Martín, Michelle A Lee, Rachel Gardner, Silvio Savarese, Jeannette Bohg, and Animesh Garg · 2019
Cited alongside, same era.
Discovering motor programs by recomposing demonstrations
Tanmay Shankar, Shubham Tulsiani, Lerrel Pinto, and Abhinav Gupta · 2019
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Cited alongside, same era.
Bellman eluder dimension: New rich classes of rl problems, and sample-efficient algorithms
Chi Jin, Qinghua Liu, and Sobhan Miryoosefi · 2021
Later among the works it cites.
Learning to simulate self-driven particles system with coordinated policy optimization
Zhenghao Peng, Quanyi Li, Ka Ming Hui, Chunxiao Liu, and Bolei Zhou · 2021
Later among the works it cites.
Learning transferable motor skills with hierarchical latent mixture policies
Dushyant Rao, Fereshteh Sadeghi, Leonard Hasenclever, Markus Wulfmeier, Martina Zambelli, Giulia Vezzani, Dhruva Tirumala, Yusuf Aytar, Josh Merel, Nicolas Heess, et al · 2021
Later among the works it cites.
Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning
Quanyi Li, Zhenghao Peng, Lan Feng, Qihang Zhang, Zhenghai Xue, and Bolei Zhou · 2022
Later among the works it cites.
Improved deep reinforcement learning with expert demonstrations for urban autonomous driving
Haochen Liu, Zhiyu Huang, and Chen Lv · 2022
Later among the works it cites.
Yiren Lu, Justin Fu, George Tucker, Xinlei Pan, Eli Bronstein, Becca Roelofs, Benjamin Sapp, Brandyn White, Aleksandra Faust, Shimon Whiteson, et al · 2022
Later among the works it cites.
Reinforcement learning with sparse rewards using guidance from offline demonstration
Desik Rengarajan, Gargi Vaidya, Akshay Sarvesh, Dileep Kalathil, and Srinivas Shakkottai · 2022
Later among the works it cites.
Tong Zhou, Letian Wang, Ruobing Chen, Wenshuo Wang, and Yu Liu · 2022
Later among the works it cites.
Safety-enhanced autonomous driving using interpretable sensor fusion transformer
Hao Shao, Letian Wang, Ruobing Chen, Hongsheng Li, and Yu Liu · 2023
Closest in time.