Fetching the paper…
Reading the bibliography…
An open problem in autonomous vehicle safety validation is building reliable models of human driving behavior in simulation.
“Estimating the dimension of a model”
Gideon Schwarz · 1978
Earlier work this paper cites.
“A behavioural car-following model for computer simulation”
Peter Gipps · 1981
Earlier work this paper cites.
“Markov games as a framework for multi-agent reinforcement learning”
Michael Littman · 1994
Earlier work this paper cites.
“Long short-term memory”
Sepp Hochreiter and J\"urgen Schmidhuber · 1997
Earlier work this paper cites.
“Is imitation learning the route to humanoid robots?”
Stefan Schaal · 1999
Earlier work this paper cites.
“Congested traffic states in empirical observations and microscopic simulations”
Martin Treiber, Ansgar Hennecke and Dirk Helbing · 2000
Earlier work this paper cites.
“Improving predictive inference under covariate shift by weighting the log-likelihood function”
Hidetoshi Shimodaira · 2000
Earlier work this paper cites.
“Develop a car-following model using data collected by five-wheel system”
Jia Hongfei, Juan Zhicai and Ni Anning · 2003
Earlier work this paper cites.
“Apprenticeship learning via inverse reinforcement learning”
Pieter Abbeel and Andrew Ng · 2004
Earlier work this paper cites.
“US highway 101 dataset”, 2007
J. Colyar and J. Halkias · 2007
Earlier work this paper cites.
“General lane-changing model MOBIL for car-following models”
Arne Kesting, Martin Treiber and Dirk Helbing · 2007
Earlier work this paper cites.
“Neural Agent Car-Following Models”
S. Panwai and H. Dia · 2007
Earlier work this paper cites.
“Signal modelling and Hidden Markov Models for driving manoeuvre recognition and driver fault diagnosis in an urban road scenario”
P. Boyraz, M. Acar and D. Kerr · 2007
Earlier work this paper cites.
“A game-theoretic approach to apprenticeship learning”
Umar Syed and Robert Schapire · 2008
Earlier work this paper cites.
“Maximum entropy inverse reinforcement learning”
Brian Ziebart, Andrew Maas, J Bagnell and Anind Dey · 2008
Earlier work this paper cites.
“Multiagent reinforcement learning for urban traffic control using coordination graphs”
Lior Kuyer, Shimon Whiteson, Bram Bakker and Nikos Vlassis · 2008
Earlier work this paper cites.
“Convolutional deep belief networks for scalable unsupervised learning of hierarchical representations”
Honglak Lee, Roger Grosse, Rajesh Ranganath and Andrew Ng · 2009
Earlier work this paper cites.
“Modeling interaction via the principle of maximum causal entropy”
Brian Ziebart, J Bagnell and Anind Dey · 2010
Earlier work this paper cites.
“Recurrent policy gradients”
Daan Wierstra, Alexander F\"orster, Jan Peters and J\"urgen Schmidhuber · 2010
Earlier work this paper cites.
“A reduction of imitation learning and structured prediction to no-regret online learning”
St\’ephane Ross, Geoffrey Gordon and Drew Bagnell · 2011
Earlier work this paper cites.
“Probabilistic trajectory prediction with gaussian mixture models”
J\"urgen Wiest, Matthias H\"offken, Ulrich Kreel and Klaus Dietmayer · 2012
Earlier work this paper cites.
“Continuous inverse optimal control with locally optimal examples”
Sergey Levine and Vladlen Koltun · 2012
Earlier work this paper cites.
“ImageNet Classification with Deep Convolutional Neural Networks”
Alex Krizhevsky, Ilya Sutskever and Geoffrey. Hinton · 2012
Earlier work this paper cites.
“A modified car-following model based on a neural network model of the human driver effects”
Alireza Khodayari, Ali Ghaffari, Reza Kazemi and Reinhard Braunstingl · 2012
Earlier work this paper cites.
“Modelling stop intersection approaches using gaussian processes”
Alexandre Armand, David Filliat and Javier Ibanez-Guzman · 2013
Cited alongside, same era.
“Infinite time horizon maximum causal entropy inverse reinforcement learning”
Michael Bloem and Nicholas Bambos · 2014
Cited alongside, same era.
“Generative adversarial nets”
Ian Goodfellow et al · 2014
Cited alongside, same era.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
“Comparison of parametric and non-parametric approaches for vehicle speed prediction”
St\’ephanie Lef\‘evre, Chao Sun, Ruzena Bajcsy and Christian Laugier · 2014
Cited alongside, same era.
“Vehicle lateral position prediction: A small step towards a comprehensive risk assessment system”
Qiang Liu, B. Lathrop and V. Butakov · 2014
“Improved training of Wasserstein GANs”
Ishaan Gulrajani et al · 2017
Later among the works it cites.
“Simultaneous policy learning and latent state inference for imitating driver behavior”
Jeremy Morton and Mykel Kochenderfer · 2017
Later among the works it cites.
“How would surround vehicles move? a unified framework for maneuver classification and motion prediction”
Nachiket Deo, Akshay Rangesh and Mohan Trivedi · 2018
Later among the works it cites.
“Burn-in demonstrations for multi-modal imitation learning”
Alex Kuefler and Mykel Kochenderfer · 2018
Later among the works it cites.
“Multi-agent imitation learning for driving simulation”
Raunak Bhattacharyya et al · 2018
Later among the works it cites.
“Multi-agent generative adversarial imitation learning”
Jiaming Song, Hongyu Ren, Dorsa Sadigh and Stefano Ermon · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation”
Kyunghyun Cho et al · 2014
Cited alongside, same era.
“Trust region policy optimization”
John Schulman et al · 2015
Cited alongside, same era.
“The multilayer perceptron approach to lateral motion prediction of surrounding vehicles for autonomous vehicles”
Seungje Yoon and Dongsuk Kum · 2016
Cited alongside, same era.
“Social LSTM: Human trajectory prediction in crowded spaces”
Alexandre Alahi et al · 2016
Cited alongside, same era.
“Analysis of recurrent neural networks for probabilistic modeling of driver behavior”
Jeremy Morton, Tim Wheeler and Mykel Kochenderfer · 2016
Cited alongside, same era.
“Generative adversarial imitation learning”
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
Later among the works it cites.
“Lenient Multi-Agent Deep Reinforcement Learning”
Gregory Palmer, Karl Tuyls, Daan Bloembergen and Rahul Savani · 2018
Later among the works it cites.
“Relational inductive biases, deep learning, and graph networks”
Peter Battaglia et al · 2018
Later among the works it cites.
“BézierVAE: Improved trajectory modeling using variational autoencoders for the safety validation of highly automated vehicles”
Robert Krajewski, Tobias Moers, Adrian Meister and Lutz Eckstein · 2019
Later among the works it cites.
“Precog: Prediction conditioned on goals in visual multi-agent settings”
Nicholas Rhinehart, Rowan McAllister, Kris Kitani and Sergey Levine · 2019
Later among the works it cites.
“Multiple futures prediction”
Charlie Tang and Russ Salakhutdinov · 2019
Later among the works it cites.
“Decision making for autonomous vehicles at unsignalized intersection in presence of malicious vehicles”
Sasinee Pruekprasert et al · 2019
Later among the works it cites.
“Interactive decision making for autonomous vehicles in dense traffic”
David Isele · 2019
Later among the works it cites.
“Hierarchical game-theoretic planning for autonomous vehicles”
Jaime Fisac et al · 2019
Later among the works it cites.
“Human-like decision-making for automated driving in highways”
David Gonz\’alez, Mario Garz\’on, Jilles Dibangoye and Christian Laugier · 2019
Later among the works it cites.
“Multi-Agent Adversarial Inverse Reinforcement Learning”
Lantao Yu, Jiaming Song and Stefano Ermon · 2019
Later among the works it cites.
“Simulating emergent properties of human driving behavior using reward augmented multi-agent imitation learning”
Raunak Bhattacharyya et al · 2019
Later among the works it cites.
“SQIL: imitation learning via regularized behavioral cloning”
Siddharth Reddy, Anca Dragan and Sergey Levine · 2019
Later among the works it cites.
“Multi-agent Adversarial Inverse Reinforcement Learning with Latent Variables”
Nate Gruver, Jiaming Song, Mykel Kochenderfer and Stefano Ermon · 2020
Closest in time.
“Parameter Sharing is Surprisingly Useful for Multi-Agent Deep Reinforcement Learning”
Justin Terry et al · 2020
Closest in time.
“Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications”
Thanh Nguyen, Ngoc Nguyen and Saeid Nahavandi · 2020
Closest in time.
“Trafficsim: Learning to simulate realistic multi-agent behaviors”
Simon Suo, Sebastian Regalado, Sergio Casas and Raquel Urtasun · 2021
Closest in time.
“Deep Implicit Coordination Graphs for Multi-agent Reinforcement Learning”
Sheng Li et al · 2021
Closest in time.