Fetching the paper…
Reading the bibliography…
Humans drive in a holistic fashion which entails, in particular, understanding dynamic road events and their evolution.
V. T. Covello and M. W. Merkhofer, “An evaluation of the state of the art,” in Risk Assessment Methods . Springer, 1993, pp. 239–265
1993
Earlier work this paper cites.
M. Bertozzi, A. Broggi, and A. Fascioli, “Vision-based intelligent vehicles: State of the art and perspectives,” Robotics and Autonomous Systems , vol. 32, no. 1, pp. 1 – 16, 2000
2000
Earlier work this paper cites.
N. Dalal and B. Triggs, “Histograms of oriented gradients for human detection,” in Computer Vision and Pattern Recognition, 2005. CVPR 2005. IEEE Computer Society Conference on , vol. 1. IEEE, 2005, pp. 886–893
2005
Earlier work this paper cites.
J. Winn and J. Shotton, “The layout consistent random field for recognizing and segmenting partially occluded objects,” in Computer Vision and Pattern Recognition, 2006 IEEE Computer Society Conference on , vol. 1. IEEE, 2006, pp. 37–44
2006
Earlier work this paper cites.
A. Lerner, Y. Chrysanthou, and D. Lischinski, “Crowds by example,” in Computer graphics forum , vol. 26, no. 3. Wiley Online Library, 2007, pp. 655–664
2007
Earlier work this paper cites.
G. J. Brostow, J. Shotton, J. Fauqueur, and R. Cipolla, “Segmentation and recognition using structure from motion point clouds,” in European conference on computer vision , 2008, pp. 44–57
2008
Earlier work this paper cites.
A. Ess, B. Leibe, K. Schindler, and L. Van Gool, “A mobile vision system for robust multi-person tracking,” in 2008 IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2008, pp. 1–8
2008
Earlier work this paper cites.
M. Enzweiler and D. M. Gavrila, “Monocular pedestrian detection: Survey and experiments,” IEEE transactions on pattern analysis and machine intelligence , vol. 31, no. 12, pp. 2179–2195, 2008
2008
Earlier work this paper cites.
C. Wojek, S. Walk, and B. Schiele, “Multi-cue onboard pedestrian detection,” in 2009 IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2009, pp. 794–801
2009
Earlier work this paper cites.
S. Pellegrini, A. Ess, and L. Van Gool, “Improving data association by joint modeling of pedestrian trajectories and groupings,” in European conference on computer vision . Springer, 2010, pp. 452–465
2010
Earlier work this paper cites.
G. Pandey, J. R. McBride, and R. M. Eustice, “Ford campus vision and lidar data set,” International Journal of Robotics Research , vol. 30, no. 13, pp. 1543–1552, 2011
2011
Earlier work this paper cites.
S. Oh, A. Hoogs, A. Perera, N. Cuntoor, C.-C. Chen, J. T. Lee, S. Mukherjee, J. Aggarwal, H. Lee, L. Davis et al. , “A large-scale benchmark dataset for event recognition in surveillance video,” in CVPR 2011 . IEEE, 2011, pp. 3153–3160
2011
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in 2012 IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2012, pp. 3354–3361
2012
Earlier work this paper cites.
K. Soomro, A. R. Zamir, and M. Shah, “Ucf101: A dataset of 101 human actions classes from videos in the wild,” 2012
2012
Earlier work this paper cites.
C. Wolf, J. Mille, E. Lombardi, O. Celiktutan, M. Jiu, M. Baccouche, E. Dellandréa, C.-E. Bichot, C. Garcia, and B. Sankur, “The LIRIS Human activities dataset and the ICPR 2012 human activities recognition and localization competition,” LIRIS UMR 5205 CNRS/INSA de Lyon/Université Claude Bernard Lyon 1/Université Lumière Lyon 2/École Centrale de Lyon, Tech. Rep., 2012. [Online]. Available: http://liris.cnrs.fr/publis/?id=5498
2012
Earlier work this paper cites.
2012
Earlier work this paper cites.
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black, “Towards understanding action recognition,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2013, pp. 3192–3199
2013
Earlier work this paper cites.
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun, “Vision meets robotics: The kitti dataset,” The International Journal of Robotics Research , vol. 32, no. 11, pp. 1231–1237, 2013
2013
Earlier work this paper cites.
J.-L. Blanco-Claraco, F.-Á. Moreno-Dueñas, and J. González-Jiménez, “The málaga urban dataset: High-rate stereo and lidar in a realistic urban scenario,” The International Journal of Robotics Research , vol. 33, no. 2, pp. 207–214, 2014
2014
Earlier work this paper cites.
Y. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar, “Thumos challenge: Action recognition with a large number of classes,” http://crcv.ucf.edu/THUMOS14 , 2014
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in Advances in neural information processing systems , 2014, pp. 568–576
2014
Earlier work this paper cites.
G. Gkioxari and J. Malik, “Finding action tubes,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition , 2015
2015
Earlier work this paper cites.
F. Caba Heilbron, V. Escorcia, B. Ghanem, and J. Carlos Niebles, “Activitynet: A large-scale video benchmark for human activity understanding,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 961–970
2015
Earlier work this paper cites.
P. Weinzaepfel, Z. Harchaoui, and C. Schmid, “Learning to track for spatio-temporal action localization,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition , June 2015
2015
Earlier work this paper cites.
M. e. a. Maurer, Autonomous driving: technical, legal and social aspects . Springer Nature, 2016
2016
Earlier work this paper cites.
A. Broggi, Alberto et al. C. Laugier, “Intelligent vehicles,” in Springer Handbook of Robotics . Springer, 2016, pp. 1627–1656
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in Proceedings of CVPR 2016 , 2016, pp. 3213–3223
2016
Earlier work this paper cites.
H. Jung, Y. Oto, O. M. Mozos, Y. Iwashita, and R. Kurazume, “Multi-modal panoramic 3d outdoor datasets for place categorization,” in 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2016, pp. 4545–4550
2016
Earlier work this paper cites.
A. Robicquet, A. Sadeghian, A. Alahi, and S. Savarese, “Learning social etiquette: Human trajectory understanding in crowded scenes,” in European conference on computer vision . Springer, 2016, pp. 549–565
2016
Earlier work this paper cites.
G. Ros, L. Sellart, J. Materzynska, D. Vazquez, and A. M. Lopez, “The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 3234–3243
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Soomro, H. Idrees, and M. Shah, “Predicting the where and what of actors and actions through online action localization,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 2648–2657
2016
Earlier work this paper cites.
X. Peng and C. Schmid, “Multi-region two-stream r-cnn for action detection,” in European Conference on Computer Vision , 2016, pp. 744–759
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
S. Saha, G. Singh, M. Sapienza, P. H. S. Torr, and F. Cuzzolin, “Deep learning for detecting multiple space-time action tubes in videos,” in British Machine Vision Conference , 2016
2016
Earlier work this paper cites.
K. Korosec, “Toyota is betting on this startup to drive its self-driving car plans forward,” Available at: http://fortune.com/2017/09/27/toyota-self-driving-car-luminar/
2017
Earlier work this paper cites.
J. Redmon and A. Farhadi, “Yolo9000: Better, faster, stronger,” in IEEE Int. Conf. on Computer Vision and Pattern Recognition , 2017, pp. 6517––6525
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
W. Maddern, G. Pascoe, C. Linegar, and P. Newman, “1 year, 1000 km: The oxford robotcar dataset,” The International Journal of Robotics Research , vol. 36, no. 1, pp. 3–15, 2017
2017
Cited alongside, same era.
G. Singh, S. Saha, M. Sapienza, P. Torr, and F. Cuzzolin, “Online real-time multiple spatiotemporal action localisation and prediction,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 3637–3646
2017
Cited alongside, same era.
G. Neuhold, T. Ollmann, S. Rota Bulo, and P. Kontschieder, “The mapillary vistas dataset for semantic understanding of street scenes,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 4990–4999
2017
Cited alongside, same era.
A. Rudenko, L. Palmieri, M. Herman, K. M. Kitani, D. M. Gavrila, and K. O. Arras, “Human motion trajectory prediction: A survey,” arXiv preprint arXiv1905.06113, 2019
2019
Later among the works it cites.
C. Feichtenhofer, H. Fan, J. Malik, and K. He, “Slowfast networks for video recognition,” in Proceedings of the IEEE international conference on computer vision , 2019, pp. 6202–6211
2019
Later among the works it cites.
P. Wang, X. Huang, X. Cheng, D. Zhou, Q. Geng, and R. Yang, “The apolloscape open dataset for autonomous driving and its application,” IEEE transactions on pattern analysis and machine intelligence , 2019
2019
Later among the works it cites.
Z. Che, G. Li, T. Li, B. Jiang, X. Shi, X. Zhang, Y. Lu, G. Wu, Y. Liu, and J. Ye, “D 2
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Zhang, R. Benenson, and B. Schiele, “Citypersons: A diverse dataset for pedestrian detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 3213–3221
2017
Cited alongside, same era.
A. Rasouli, I. Kotseruba, and J. K. Tsotsos, “Are they going to cross? a benchmark dataset and baseline for pedestrian crosswalk behavior,” in Proceedings of the IEEE International Conference on Computer Vision Workshops , 2017, pp. 206–213
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
R. Goyal, S. E. Kahou, V. Michalski, J. Materzyńska, S. Westphal, H. Kim, V. Haenel, I. Fruend, P. Yianilos, M. Mueller-Freitag, F. Hoppe, C. Thurau, I. Bax, and R. Memisevic, “The ”something something” video database for learning and evaluating visual common sense,” 2017
2017
Cited alongside, same era.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? a new model and the kinetics dataset,” in IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 4724–4733
2017
Cited alongside, same era.
2017
Cited alongside, same era.
V. Kalogeiton, P. Weinzaepfel, V. Ferrari, and C. Schmid, “Action tubelet detector for spatio-temporal action localization,” in Proc. Int. Conf. Computer Vision , 2017
2017
Cited alongside, same era.
A. Patil, S. Malla, H. Gang, and Y.-T. Chen, “The h3d dataset for full-surround 3d multi-object detection and tracking in crowded urban scenes,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 9552–9557
2019
Later among the works it cites.
J. Chang, Ming-Fang D. Wang, P. Carr, S. Lucey, D. Ramanan, and others et al., “Argoverse: 3d tracking and forecasting with rich maps,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 8748–8757
2019
Later among the works it cites.
R. Kesten, M. Usman, J. Houston, T. Pandya, K. Nadhamuni, A. Ferreira, M. Yuan, B. Low, A. Jain, P. Ondruska et al. , “Lyft level 5 av dataset 2019,” urlhttps://level5. lyft. com/dataset , 2019
2019
Later among the works it cites.
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine, V. Vasudevan, W. Han, J. Ngiam, H. Zhao, A. Timofeev, S. Ettinger, M. Krivokon, A. Gao, A. Joshi, Y. Zhang, J. Shlens, Z. Chen, and D. Anguelov, “Scalability in perception for autonomous driving: Waymo open dataset,” 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
J. Geyer, Y. Kassahun, M. Mahmudi, X. Ricou, R. Durgesh, A. S. Chung, L. Hauswald, V. H. Pham, M. Mühlegg, S. Dorn et al. , “A2d2: Aev autonomous driving dataset,” Note: http://www. a2d2. audi Cited by , vol. 1, no. 4, 2019
2019
Later among the works it cites.
Y. Yao, M. Xu, C. Choi, D. J. Crandall, E. M. Atkins, and B. Dariush, “Egocentric vision-based future vehicle localization for intelligent driving assistance systems,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 9711–9717
2019
Later among the works it cites.
R. Chandra, U. Bhattacharya, A. Bera, and D. Manocha, “Traphic: Trajectory prediction in dense and heterogeneous traffic using weighted interactions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 8483–8492
2019
Later among the works it cites.
A. Rasouli, I. Kotseruba, T. Kunic, and J. K. Tsotsos, “Pie: A large-scale dataset and models for pedestrian intention estimation and trajectory prediction,” in Proceedings of the IEEE International Conference on Computer Vision , 2019, pp. 6262–6271
2019
Later among the works it cites.
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall, “Semantickitti: A dataset for semantic scene understanding of lidar sequences,” 2019
2019
Later among the works it cites.
M. Monfort, A. Andonian, B. Zhou, K. Ramakrishnan, S. A. Bargal, T. Yan, L. Brown, Q. Fan, D. Gutfruend, C. Vondrick et al. , “Moments in time dataset: one million videos for event understanding,” IEEE Transactions on Pattern Analysis and Machine Intelligence , pp. 1–8, 2019
2019
Later among the works it cites.
C.-Y. Wu, C. Feichtenhofer, H. Fan, K. He, P. Krahenbuhl, and R. Girshick, “Long-term feature banks for detailed video understanding,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 284–293
2019
Later among the works it cites.
X. Yang, X. Yang, M.-Y. Liu, F. Xiao, L. S. Davis, and J. Kautz, “Step: Spatio-temporal progressive learning for video action detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 264–272
2019
Later among the works it cites.
J. Zhao and C. G. Snoek, “Dance with flow: Two-in-one stream action detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 9935–9944
2019
Later among the works it cites.
L. Song, S. Zhang, G. Yu, and H. Sun, “Tacnet: Transition-aware context network for spatio-temporal action detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019, pp. 11 987–11 995
2019
Later among the works it cites.
T. Kong, F. Sun, H. Liu, Y. Jiang, and J. Shi, “Consistent optimization for single-shot object detection,” 2019
2019
Later among the works it cites.
G. SingH and F. Cuzzolin, “Recurrent convolutions for causal 3d cnns,” in Proceedings of the IEEE International Conference on Computer Vision Workshops , 2019, pp. 0–0
2019
Later among the works it cites.
G. I. Parisi, R. Kemker, J. L. Part, C. Kanan, and S. Wermter, “Continual lifelong learning with neural networks: A review,” Neural Networks , vol. 113, pp. 54–71, 2019
2019
Later among the works it cites.
F. Cuzzolin, A. Morelli, B. Cirstea, and B. J. Sahakian, “Knowing me, knowing you: Theory of mind in AI,” Psychological Medicine , vol. 50, no. 7, pp. 1057–1061, May 2020
2020
Later among the works it cites.
A. Rasouli and J. K. Tsotsos, “Autonomous vehicles that interact with pedestrians: A survey of theory and practice,” IEEE Transactions on Intelligent Transportation Systems , vol. 21, no. 3, pp. 900––918, 2020
2020
Later among the works it cites.
L. Ding, J. Terwilliger, R. Sherony, B. Reimer, and L. Fridman, “MIT DriveSeg (Manual) Dataset,” 2020
2020
Later among the works it cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 11 621–11 631
2020
Later among the works it cites.
S. Malla, B. Dariush, and C. Choi, “Titan: Future forecast using action priors,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 11 186–11 196
2020
Later among the works it cites.
C. Feichtenhofer, “X3d: Expanding architectures for efficient video recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 203–213
2020
Later among the works it cites.
Y. Li, Z. Wang, L. Wang, and G. Wu, “Actions as moving points,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020
2020
Later among the works it cites.
J. Tang, J. Xia, X. Mu, B. Pang, and C. Lu, “Asynchronous interaction aggregation for action detection,” in European Conference on Computer Vision . Springer, 2020, pp. 71–87
2020
Later among the works it cites.
M. Li, Y.-X. Wang, and D. Ramanan, “Towards streaming perception,” in European Conference on Computer Vision . Springer, 2020, pp. 473–488
2020
Later among the works it cites.
W. Liu, G. Kang, P.-Y. Huang, X. Chang, Y. Qian, J. Liang, L. Gui, J. Wen, and P. Chen, “Argus: Efficient activity detection system for extended video analysis,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision Workshops , 2020, pp. 126–133
2020
Later among the works it cites.
P. Sun, H. Kretzschmar, X. Dotiwalla, A. Chouard, V. Patnaik, P. Tsui, J. Guo, Y. Zhou, Y. Chai, B. Caine et al. , “Scalability in perception for autonomous driving: Waymo open dataset,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 2446–2454
2020
Later among the works it cites.
Y. Liao, J. Xie, and A. Geiger, “KITTI-360: A novel dataset and benchmarks for urban scene understanding in 2d and 3d,” arXiv.org , vol. 2109.13410, 2021
2021
Closest in time.
2021
Closest in time.
J. Pan, S. Chen, M. Z. Shou, Y. Liu, J. Shao, and H. Li, “Actor-context-actor relation network for spatio-temporal action localization,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 464–474
2021
Closest in time.
2021
Closest in time.