Fetching the paper…
Reading the bibliography…
In many settings where multiple agents interact, the optimal choices for each agent depend heavily on the choices of the others.
On the Stackelberg Strategy in Nonzero-sum Games
Marwaan Simaan and Jose B Cruz. 1973 · 1973
Earlier work this paper cites.
The Complexity of Markov Decision Processes
Christos H. Papadimitriou and John N. Tsitsiklis. 1987 · 1987
Earlier work this paper cites.
Learning policies for partially observable environments: Scaling up. In International Conference on Machine Learning (ICML)
Michael L Littman, Anthony R Cassandra, and Leslie Pack Kaelbling. 1995 · 1995
Earlier work this paper cites.
Planning and Acting in Partially Observable Stochastic Domains
Leslie Pack Kaelbling, Michael L. Littman, and Anthony R. Cassandra. 1998 · 1998
Earlier work this paper cites.
Congested Traffic States in Empirical Observations and Microscopic Simulations
Martin Treiber, Ansgar Hennecke, and Dirk Helbing. 2000 · 2000
Earlier work this paper cites.
Complexity Results About Nash Equilibria
Vincent Conitzer and Tuomas Sandholm. 2002 · 2002
Earlier work this paper cites.
Iterative Linear Quadratic Regulator Design for Nonlinear Biological Movement Systems. In ICINCO (1)
Weiwei Li and Emanuel Todorov. 2004 · 2004
Earlier work this paper cites.
A Generalized Iterative LQG Method for Locally-Optimal Feedback Control of Constrained Nonlinear Stochastic Systems. In IEEE American Control Conference (ACC)
Emanuel Todorov and Weiwei Li. 2005 · 2005
Earlier work this paper cites.
Goal Inference as Inverse Planning. In Proceedings of the Annual Meeting of the Cognitive Science Society
Chris L Baker, Joshua B Tenenbaum, and Rebecca R Saxe. 2007 · 2007
Earlier work this paper cites.
General Lane-Changing Model MOBIL for Car-Following Models
Arne Kesting, Martin Treiber, and Dirk Helbing. 2007 · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning.. In AAAI Conference on Artificial Intelligence (AAAI)
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey. 2008 · 2008
Earlier work this paper cites.
Julia: A Fast Dynamic Language for Technical Computing
Jeff Bezanson, Stefan Karpinski, Viral B. Shah, and Alan Edelman. 2012 · 2012
Earlier work this paper cites.
Stackelberg Game Based Model of Highway Driving. In ASME Dynamic Systems and Control Conference joint with the JSME Motion and Vibration Conference
Je Hong Yoo and Reza Langari. 2012 · 2012
Cited alongside, same era.
Safe Sequential Path Planning of Multi-Vehicle Systems via Double-Obstacle Hamilton-Jacobi-Isaacs Variational Inequality. In European Control Conference (ECC)
Mo Chen, Jaime F Fisac, Shankar Sastry, and Claire J Tomlin. 2015 · 2015
Cited alongside, same era.
MPDM: Multipolicy Decision-Making in Dynamic, Uncertain Environments for Autonomous Driving. In IEEE International Conference on Robotics and Automation (ICRA)
Alexander G Cunningham, Enric Galceran, Ryan M Eustice, and Edwin Olson. 2015 · 2015
Cited alongside, same era.
Multipolicy Decision-Making for Autonomous Driving via Changepoint-based Behavior Prediction.. In Robotics: Science and Systems (RSS)
Enric Galceran, Alexander G Cunningham, Ryan M Eustice, and Edwin Olson. 2015 · 2015
Cited alongside, same era.
Social GAN: Socially Acceptable Trajectories with Generative Adversarial Networks. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Agrim Gupta, Justin Johnson, Li Fei-Fei, Silvio Savarese, and Alexandre Alahi. 2018 · 2018
Later among the works it cites.
A Belief State Planner for Interactive Merge Maneuvers in Congested Traffic. In IEEE International Conference on Intelligent Transportation Systems (ITSC)
Constantin Hubmann, Jens Schulz, Gavin Xu, Daniel Althoff, and Christoph Stiller. 2018 · 2018
Later among the works it cites.
On the Convergence of Gradient-Based Learning in Continuous Games
Eric Mazumdar and Lillian J Ratliff. 2018 · 2018
Later among the works it cites.
Multimodal Probabilistic Model-Based Planning for Human-Robot Interaction. In IEEE International Conference on Robotics and Automation (ICRA)
Edward Schmerling, Karen Leung, Wolf Vollprecht, and Marco Pavone. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mykel J. Kochenderfer. 2015 · 2015
Cited alongside, same era.
Autonomous Navigation in Dynamic Social Environments Using Multi-Policy Decision Making. In IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
Dhanvin Mehta, Gonzalo Ferrer, and Edwin Olson. 2016 · 2016
Cited alongside, same era.
Predicting Actions to Act Predictably: Cooperative Partial Motion Planning with Maximum Entropy Models. In IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
Mark Pfeiffer, Ulrich Schwesinger, Hannes Sommer, Enric Galceran, and Roland Siegwart. 2016 · 2016
Cited alongside, same era.
On the Characterization of Local Nash Equilibria in Continuous Games
Lillian J Ratliff, Samuel A Burden, and S Shankar Sastry. 2016 · 2016
Cited alongside, same era.
Explainable Agency for Intelligent Autonomous Systems. In Twenty-Ninth IAAI Conference
Pat Langley, Ben Meadows, Mohan Sridharan, and Dongkyu Choi. 2017 · 2017
Cited alongside, same era.
The Value of Inferring the Internal State of Traffic Participants for Autonomous Freeway Driving. In IEEE American Control Conference (ACC)
Zachary N Sunberg, Christopher J Ho, and Mykel J. Kochenderfer. 2017 · 2017
Cited alongside, same era.
Probabilistic Model for Interaction Aware Planning in Merge Scenarios
Erik Ward, Niclas Evestedt, Daniel Axehill, and John Folkesson. 2017 · 2017
Cited alongside, same era.
Planning for Autonomous Cars that Leverage Effects on Human Actions.. In Robotics: Science and Systems (RSS)
Dorsa Sadigh, Shankar Sastry, Sanjit A Seshia, and Anca D Dragan. 2016a
Cited in the paper.
A Scalable Framework for Real-Time Multi-Robot, Multi-Human Collision Avoidance. In IEEE International Conference on Robotics and Automation (ICRA)
Andrea Bajcsy, Sylvia L Herbert, David Fridovich-Keil, Jaime F Fisac, Sampada Deglurkar, Anca D Dragan, and Claire J Tomlin. 2019 · 2019
Later among the works it cites.
Cooperation-Aware Reinforcement Learning for Merging in Dense Traffic
Maxime Bouton, Alireza Nakhaei, Kikuo Fujimura, and Mykel J. Kochenderfer. 2019 · 2019
Later among the works it cites.
Hierarchical Game-Theoretic Planning for Autonomous Vehicles. In IEEE International Conference on Robotics and Automation (ICRA)
Jaime F Fisac, Eli Bronstein, Elis Stefansson, Dorsa Sadigh, S Shankar Sastry, and Anca D Dragan. 2019 · 2019
Later among the works it cites.
Confidence-Aware Motion Prediction for Real-Time Collision Avoidance
David Fridovich-Keil, Andrea Bajcsy, Jaime F Fisac, Sylvia L Herbert, Steven Wang, Anca D Dragan, and Claire J Tomlin. 2019a · 2019
Later among the works it cites.
David Fridovich-Keil, Ellis Ratner, Lasse Peters, Anca D. Dragan, and Claire J. Tomlin. 2019b · 2019
Later among the works it cites.
An Iterative Quadratic Method for General-Sum Differential Games with Feedback Linearizable Dynamics
David Fridovich-Keil, Vicenc Rubies-Royo, and Claire J Tomlin. 2019c · 2019
Later among the works it cites.
On Finding Local Nash Equilibria (and Only Local Nash Equilibria) in Zero-Sum Games
Eric V Mazumdar, Michael I Jordan, and S Shankar Sastry. 2019 · 2019
Later among the works it cites.