Fetching the paper…
Reading the bibliography…
Driving Scene understanding is a key ingredient for intelligent transportation systems.
A Critical View of Driver Behavior Models: What Do We Know, What Should We Do?
J. A. Michon · 1985
Earlier work this paper cites.
ImageNet: A large-scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L. jia Li, K. Li, and L. Fei-fei · 2009
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2009
Earlier work this paper cites.
Are We Ready for Autonomous Driving? The KITTI Vision Benchmark Suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Activity forecasting
K. Kitani, B. Ziebart, J. Bagnell, and M. Hebert · 2012
Earlier work this paper cites.
Will the Pedestrian Cross? A Study on Pedestrian Path Prediction
C. G. Keller and D. M. Gavrila · 2014
Earlier work this paper cites.
Context-based Pedestrian Path Prediction
J. Kooij, N. Schneider, F. Flohr, and D. Gavrila · 2014
Earlier work this paper cites.
A Novel Goal Oriented Concept for Situation Representation for ADAS and Automated Driving
M. Schmidt, U. Hofmann, and M. Bouzouraa · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
A. Agrawal, J. Lu, S. Antol, M. Mitchell, C. L. Zitnick, D. Batra, and D. Parikh · 2015
Earlier work this paper cites.
DeepDriving: Learning Affordance for Direct Perception in Autonomous Driving
C. Chen, A. Seff, A. Kornhauser, and J. Xiao · 2015
Earlier work this paper cites.
ActivityNet: A Large-Scale Video Benchmark for Human Activity Understanding
F. C. Heilbron, V. Escorcia, B. Ghanem, and J. C. Niebles · 2015
Earlier work this paper cites.
Car that Knows Before You Do: Anticipating Maneuvers via Learning Temporal Driving Models
A. Jain, H. S. Koppula, B. Raghavan, S. Soh, and A. Saxena · 2015
Cited alongside, same era.
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
Learning to Compose Neural Networks for Question Answering
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Cited alongside, same era.
The Cityscapes Dataset for Semantic Urban Scene Understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Cited alongside, same era.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
J. Johnson, B. Hariharan, L. van der Maaten, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Later among the works it cites.
The Kinetics Human Action Video Dataset
W. Kay, J. Carreira, K. Simonyan, B. Zhang, C. Hillier, S. Vijayanarasimhan, F. Viola, T. Green, T. Back, P. Natsev, M. Suleyman, and A. Zisserman · 2017
Later among the works it cites.
DESIRE: Distant Future Prediction in Dynamic Scenes with Interacting Agents
N. Lee, W. Choi, P. Vernaza, C. B. Choy, P. H. S. Torr, and M. Chandraker · 2017
Later among the works it cites.
Focal Loss for Dense Object Detection
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollar · 2017
Later among the works it cites.
1 Year, 1000km: The Oxford RobotCar Dataset
W. Maddern, G. Pascoe, C. Linegar, and P. Newman · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Santana and G. Hotz · 2016
Cited alongside, same era.
Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding
G. A. Sigurdsson, G. Varol, X. Wang, A. Farhadi, I. Laptev, and A. Gupta · 2016
Cited alongside, same era.
Inception-v4, inception-resnet and the impact of residual connections on learning
C. Szegedy, S. Ioffe, and V. Vanhoucke · 2016
Cited alongside, same era.
End-To-End Learning of Driving Models From Large-Scale Video Datasets
H. Xu, Y. Gao, F. Yu, and T. Darrell · 2016
Cited alongside, same era.
Mask R-CNN
K. He, G. Gkioxari, P. Dollar, and R. Girshick · 2017
Cited alongside, same era.
Learning to Reason: End-To-End Module Networks for Visual Question Answering
R. Hu, J. Andreas, M. Rohrbach, T. Darrell, and K. Saenko · 2017
Cited alongside, same era.
Learning Image Representations Tied to Egomotion from Unlabeled Video
D. Jayaraman and K. Grauman · 2017
Cited alongside, same era.
The BDD-Nexar Collective: A Large-Scale, Crowsourced, Dataset of Driving Scenes
V. Madhavan and T. Darrell · 2017
Later among the works it cites.
Jointly Learning Energy Expenditures and Activities using Egocentric Multimodal Signals
K. Nakamura, S. Yeung, A. Alahi, and L. Fei-Fei · 2017
Later among the works it cites.
Are They Going to Cross? A Benchmark Dataset and Baseline for Pedestrian Crosswalk Behavior
A. Rasouli, I. Kotseruba, and J. K. Tsotsos · 2017
Later among the works it cites.
CDC: Convolutional-De-Convolutional Networks for Precise Temporal Action Localization in Untrimmed Videos
Z. Shou, J. Chan, A. Zareian, K. Miyazawa, and S. Chang · 2017
Later among the works it cites.
Public driving dataset
Udacity · 2017
Later among the works it cites.
Pyramid Scene Parsing Network
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia · 2017
Later among the works it cites.