Fetching the paper…
Reading the bibliography…
Out-of-training-distribution (OOD) scenarios are a common challenge of learning agents at deployment, typically leading to arbitrary deductions and poorly-informed decisions.
Contributions to the theory of statistical estimation and testing hypotheses
Wald, A · 1939
Earlier work this paper cites.
Pattern-recognizing control systems, 1964
Widrow, B. and Smith, F. W · 1964
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
Pomerleau, D. A · 1989
Earlier work this paper cites.
Brake reaction times of unalerted drivers
Taoka, G. T · 1989
Earlier work this paper cites.
Mixture density networks
Bishop, C. M · 1994
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
French, R. M · 1999
Earlier work this paper cites.
Prediction, learning, and games
Cesa-Bianchi, N. and Lugosi, G · 2006
Earlier work this paper cites.
Watch, try, learn: Meta-learning from demonstrations and reward
Zhou, A., Jang, E., Kappler, D., Herzog, A., Khansari, M., Wohlhart, P., Bai, Y., Kalakrishnan, M., Levine, S., and Finn, C · 2006
Earlier work this paper cites.
Pre-crash scenario typology for crash avoidance research, 2007
National Highway Traffic Safety Administration · 2007
Earlier work this paper cites.
Driver reaction times to familiar, but unexpected events
Coley, G., Wesley, A., Reed, N., and Parry, I · 2009
Earlier work this paper cites.
Standing balance control using a trajectory library
Liu, C. and Atkeson, C. G · 2009
Earlier work this paper cites.
Dataset shift in machine learning
Quionero-Candela, J., Sugiyama, M., Schwaighofer, A., and Lawrence, N. D · 2009
Earlier work this paper cites.
Using bisimulation for policy transfer in MDPs
Castro, P. S. and Precup, D · 2010
Earlier work this paper cites.
Practical variational inference for neural networks
Graves, A · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D · 2011
Earlier work this paper cites.
Bayesian reasoning and machine learning
Barber, D · 2012
Earlier work this paper cites.
Bayesian learning for neural networks , volume 118
Neal, R. M · 2012
Earlier work this paper cites.
Machine learning in non-stationary environments: Introduction to covariate shift adaptation
Sugiyama, M. and Kawanabe, M · 2012
Earlier work this paper cites.
Modelling extremal events: for insurance and finance , volume 33
Embrechts, P., Klüppelberg, C., and Mikosch, T · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Earlier work this paper cites.
Adaptive control processes: a guided tour
Bellman, R. E · 2015
Earlier work this paper cites.
Weight uncertainty in neural networks
Blundell, C., Cornebise, J., Kavukcuoglu, K., and Wierstra, D · 2015
Cited alongside, same era.
Probabilistic backpropagation for scalable learning of Bayesian neural networks
Hernández-Lobato, J. M. and Adams, R · 2015
Cited alongside, same era.
Variational inference with normalizing flows
Rezende, D. J. and Mohamed, S · 2015
Cited alongside, same era.
Concrete problems in AI safety
Amodei, D., Olah, C., Steinhardt, J., Christiano, P., Schulman, J., and Mané, D · 2016
Cited alongside, same era.
Dropout as a Bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Cited alongside, same era.
Learning dexterous in-hand manipulation
OpenAI, M. A., Baker, B., Chociej, M., Józefowicz, R., McGrew, B., Pachocki, J., Petron, A., Plappert, M., Powell, G., Ray, A., et al · 2018
Later among the works it cites.
R2P2: A reparameterized pushforward policy for diverse, precise generative path forecasting
Rhinehart, N., Kitani, K. M., and Vernaza, P · 2018
Later among the works it cites.
Conditional affordance learning for driving in urban environments
Sauer, A., Savinov, N., and Geiger, A · 2018
Later among the works it cites.
Solving Rubik’s cube with a robot hand
Akkaya, I., Andrychowicz, M., Chociej, M., Litwin, M., McGrew, B., Petron, A., Paino, A., Plappert, M., Powell, G., Ribas, R., et al · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Epopt: Learning robust neural network policies using model ensembles
Rajeswaran, A., Ghotra, S., Ravindran, B., and Levine, S · 2016
Cited alongside, same era.
Cad2rl: Real single-image flight without a single real image
Sadeghi, F. and Levine, S · 2016
Cited alongside, same era.
Neural autoregressive distribution estimation
Uria, B., Côté, M.-A., Gregor, K., Murray, I., and Larochelle, H · 2016
Cited alongside, same era.
Query-efficient imitation learning for end-to-end autonomous driving
Zhang, J. and Cho, K · 2016
Cited alongside, same era.
Deep reinforcement learning from human preferences
Christiano, P. F., Leike, J., Brown, T., Martic, M., Legg, S., and Amodei, D · 2017
Cited alongside, same era.
CARLA: An open urban driving simulator
Dosovitskiy, A., Ros, G., Codevilla, F., Lopez, A., and Koltun, V · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Caesar, H., Bankiti, V., Lang, A. H., Vora, S., Liong, V. E., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., and Beijbom, O · 2019
Later among the works it cites.
Multipath: Multiple probabilistic anchor trajectory hypotheses for behavior prediction
Chai, Y., Sapp, B., Bansal, M., and Anguelov, D · 2019
Later among the works it cites.
Chen, D., Zhou, B., Koltun, V., and Krähenbühl, P · 2019
Later among the works it cites.
Exploring the limitations of behavior cloning for autonomous driving
Codevilla, F., Santana, E., López, A. M., and Gaidon, A · 2019
Later among the works it cites.
Multimodal trajectory predictions for autonomous driving using deep convolutional networks
Cui, H., Radosavljevic, V., Chou, F.-C., Lin, T.-H., Nguyen, T., Huang, T.-K., Schneider, J., and Djuric, N · 2019
Later among the works it cites.
Causal confusion in imitation learning
de Haan, P., Jayaraman, D., and Levine, S · 2019
Later among the works it cites.
Model based planning with energy based models
Du, Y., Lin, T., and Mordatch, I · 2019
Later among the works it cites.
Generalizing from a few environments in safety-critical reinforcement learning
Kenton, Z., Filos, A., Evans, O., and Gal, Y · 2019
Later among the works it cites.
Lyft level 5 av dataset 2019, 2019
Kesten, R., Usman, M., Houston, J., Pandya, T., Nadhamuni, K., Ferreira, A., Yuan, M., Low, B., Jain, A., Ondruska, P., Omari, S., Shah, S., Kulkarni, A., Kazakova, A., Tao, C., Platinsky, L., Jiang, W., and Shet, V · 2019
Later among the works it cites.
Robustness to out-of-distribution inputs via task-aware generative uncertainty
McAllister, R., Kahn, G., Clune, J., and Levine, S · 2019
Later among the works it cites.
Covernet: Multimodal behavior prediction using trajectory sets
Phan-Minh, T., Grigore, E. C., Boulton, F. A., Beijbom, O., and Wolff, E. M · 2019
Later among the works it cites.
PRECOG: Prediction conditioned on goals in visual multi-agent settings
Rhinehart, N., McAllister, R., Kitani, K., and Levine, S · 2019
Later among the works it cites.
CARLA challenge, 2019
Ros, G., Koltun, V., Codevilla, F., and Lopez, M. A · 2019
Later among the works it cites.
Can you trust your model’s uncertainty? evaluating predictive uncertainty under dataset shift
Snoek, J., Ovadia, Y., Fertig, E., Lakshminarayanan, B., Nowozin, S., Sculley, D., Dillon, J., Ren, J., and Nado, Z · 2019
Later among the works it cites.
Scalability in perception for autonomous driving: An open dataset benchmark
Sun, P., Kretzschmar, H., Dotiwalla, X., Chouard, A., Patnaik, V., Tsui, P., Guo, J., Zhou, Y., Chai, Y., Caine, B., et al · 2019
Later among the works it cites.
Tang, Y. C., Zhang, J., and Salakhutdinov, R · 2019
Later among the works it cites.
Deep imitative models for flexible inference, planning, and control
Rhinehart, N., McAllister, R., and Levine, S · 2020
Closest in time.