D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in Advances in neural information processing systems , 1989, pp. 305–313
1989
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in Proceedings of the tenth international conference on machine learning , 1993, pp. 330–337
1993
Earlier work this paper cites.
A. Y. Ng, et al. , “Algorithms for inverse reinforcement learning.” in Icml , 2000, pp. 663–670
2000
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in Proceedings of the twenty-first international conference on Machine learning . ACM, 2004, p. 1
2004
Earlier work this paper cites.
J. Colyar and J. Halkias, “Us highway 101 dataset,” Federal Highway Administration (FHWA), Tech. Rep. FHWA-HRT-07-030 , 2007
2007
Earlier work this paper cites.
B. D. Ziebart, et al. , “Maximum entropy inverse reinforcement learning.” in AAAI , vol. 8. Chicago, IL, USA, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
B. D. Ziebart, et al. , “Navigate like a cabbie: Probabilistic reasoning from observed context-aware behavior,” in Proceedings of the 10th international conference on Ubiquitous computing . ACM, 2008, pp. 322–331
2008
Earlier work this paper cites.
D. G. R. Bradski and A. Kaehler, Learning Opencv, 1st Edition , 1st ed. O’Reilly Media, Inc., 2008
2008
Earlier work this paper cites.
B. D. Argall, et al. , “A survey of robot learning from demonstration,” Robotics and autonomous systems , vol. 57, no. 5, pp. 469–483, 2009
2009
Earlier work this paper cites.
P. Henry, et al. , “Learning to navigate through crowded environments,” in Robotics and Automation (ICRA), 2010 IEEE International Conference on . IEEE, 2010, pp. 981–986
2010
Earlier work this paper cites.
S. Ross and D. Bagnell, “Efficient reductions for imitation learning,” in Proceedings of the thirteenth international conference on artificial intelligence and statistics , 2010, pp. 661–668
2010
Earlier work this paper cites.
S. Ross, et al. , “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics , 2011, pp. 627–635
2011
Earlier work this paper cites.
D. Vasquez, et al. , “Inverse reinforcement learning algorithms and features for robot navigation in crowds: an experimental comparison,” in IEEE-RSJ Int. Conf. on Intelligent Robots and Systems , 2014, pp. 1341–1346
2014
Earlier work this paper cites.
I. Goodfellow, et al. , “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
T.-Y. Lin, et al. , “Microsoft coco: Common objects in context,” in k , 2014
2014
Earlier work this paper cites.
R. Girshick, “Fast R-CNN,” in Proceedings of the International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
S. Ren, et al. , “Faster R-CNN: Towards real-time object detection with region proposal networks,” in Neural Information Processing Systems (NIPS) , 2015
2015
Earlier work this paper cites.
J. Schulman, et al. , “Trust region policy optimization,” in International Conference on Machine Learning , 2015, pp. 1889–1897
2015
Earlier work this paper cites.