Fetching the paper…
Reading the bibliography…
A general-purpose intelligent robot must be able to learn autonomously and be able to accomplish multiple tasks in order to be deployed in the real world.
Application of intelligent automata to reconnaissance
C. Rosen and N. Nilsson · 1968
Earlier work this paper cites.
Vision and navigation for the Carnegie-Mellon Navlab
C. Thorpe, M. H. Hebert, T. Kanade, and S. A. Shafer · 1988
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. Pomerleau · 1989
Earlier work this paper cites.
Q-Learning
C. Watkins and P. Dayan · 1992
Earlier work this paper cites.
The cross-entropy method for combinatorial and continuous optimization
R. Rubinstein · 1999
Earlier work this paper cites.
The panda3d graphics engine
M. Goslin and M. R. Mine · 2004
Earlier work this paper cites.
Off-road obstacle avoidance through end-to-end learning
U. Muller, J. Ben, E. Cosatto, B. Flepp, and Y. LeCun · 2006
Earlier work this paper cites.
Multi-task reinforcement learning: a hierarchical Bayesian approach
A. Wilson, A. Fern, S. Ray, and P. Tadepalli · 2007
Earlier work this paper cites.
Autonomous driving in urban environments: Boss and the urban challenge
C. Urmson and et. al · 2008
Earlier work this paper cites.
Learning long-range vision for autonomous off-road driving
R. Hadsell, P. Sermanet, J. Ben, A. Erkan, M. Scoffier, K. Kavukcuoglu, U. Muller, and Y. LeCun · 2009
Earlier work this paper cites.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
R. S. Sutton, J. Modayil, M. Delp, T. Degris, P. M. Pilarski, A. White, and D. Precup · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. Gordon, and D. Bagnell · 2011
Earlier work this paper cites.
Model predictive control
E. F. Camacho and C. B. Alba · 2013
Earlier work this paper cites.
Bullet physics library
E. Coumans et al · 2013
Cited alongside, same era.
Universal value function approximators
T. Schaul, D. Horgan, K. Gregor, and D. Silver · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Deep multi-scale video prediction beyond mean square error
M. Mathieu, C. Couprie, and Y. LeCun · 2016
Cited alongside, same era.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Cited alongside, same era.
Anticipating the future by watching unlabeled video
C. Vondrick, H. Pirsiavash, and A. Torralba · 2016
Unsupervised learning of disentangled representations from video
E. Denton and V. Birodkar · 2017
Later among the works it cites.
Find your own way: Weakly-supervised segmentation of path proposals for urban autonomy
D. Barnes, W. Maddern, and I. Posner · 2017
Later among the works it cites.
Safe Visual Navigation via Deep Learning and Novelty Detection
C. Richter and N. Roy · 2017
Later among the works it cites.
Socially Aware Motion Planning with Deep Reinforcement Learning
Y. F. Chen, M. Everett, M. Liu, and J. How · 2017
Later among the works it cites.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, D. Kumaran, and R. Hadsell · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“What happens if…” Learning to Predict the Effect of Forces in Images
R. Mottaghi, M. Rastegari, A. Gupta, and A. Farhadi · 2016
Cited alongside, same era.
Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours
L. Pinto and A. Gupta · 2016
Cited alongside, same era.
Watch this: Scalable cost-function learning for path planning in urban environments
M. Wulfmeier, D. Z. Wang, and I. Posner · 2016
Cited alongside, same era.
End to end learning for self-driving cars
M. Bojarski, D. Del Testa, D. Dworakowski, B. Firner, B. Flepp, P. Goyal, L. D. Jackel, M. Monfort, U. Muller, J. Zhang, et al · 2016
Cited alongside, same era.
On multiplicative integration with recurrent neural networks
Y. Wu, S. Zhang, Y. Zhang, Y. Bengio, and R. Salakhutdinov · 2016
Cited alongside, same era.
Successor features for transfer in reinforcement learning
A. Barreto, W. Dabney, R. Munos, J. J. Hunt, T. Schaul, H. P. van Hasselt, and D. Silver · 2017
Cited alongside, same era.
F. Sadeghi and S. Levine · 2017
Later among the works it cites.
Hindsight experience replay
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. McGrew, J. Tobin, P. Abbeel, and W. Zaremba · 2017
Later among the works it cites.
CARLA: An Open Urban Driving Simulator
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun · 2017
Later among the works it cites.
End-to-end Driving via Conditional Imitation Learning
F. Codevilla, M. Müller, A. López, V. Koltun, and A. Dosovitskiy · 2018
Closest in time.
Deep Object-Centric Representations for Generalizable Robot Learning
C. Devin, P. Abbeel, T. Darrell, and S. Levine · 2018
Closest in time.
Self-supervised Deep Reinforcement Learning with Generalized Computation Graphs for Robot Navigation
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine · 2018
Closest in time.
Composable Deep Reinforcement Learning for Robotic Manipulation
T. Haarnoja, V. Pong, A. Zhou, M. Dalal, P. Abbeel, and S. Levine · 2018
Closest in time.