Fetching the paper…
Reading the bibliography…
A vehicle driving along the road is surrounded by many objects, but only a small subset of them influence the driver's decisions and actions.
M. Ma, H. Fan, and K. M. Kitani, “Going deeper into first-person activity recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 1894–1903
1903
Earlier work this paper cites.
G. Bertasius, H. Soo Park, S. X. Yu, and J. Shi, “Unsupervised learning of important objects from first-person videos,” in IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 1956–1964
1964
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
M. C. Mozer and M. Sitton, “Computational modeling of spatial attention,” Attention , vol. 9, pp. 341–393, 1998
1998
Earlier work this paper cites.
L. Itti, C. Koch, and E. Niebur, “A model of saliency-based visual attention for rapid scene analysis,” IEEE Transactions on Pattern Analysis and Machine Intelligence (PAMI) , vol. 20, no. 11, pp. 1254–1259, 1998
1998
Earlier work this paper cites.
M. Hayhoe and D. Ballard, “Eye movements in natural behavior,” Trends in cognitive sciences , vol. 9, no. 4, pp. 188–194, 2005
2005
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman, “The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results,” http://www.pascal-network.org/challenges/VOC/voc2007/workshop/index.html
2007
Earlier work this paper cites.
S. Perone, K. L. Madole, S. Ross-Sheehy, M. Carey, and L. M. Oakes, “The relation between infants’ activity with objects and attention to object appearance.” Developmental psychology , vol. 44, no. 5, p. 1242, 2008
2008
Earlier work this paper cites.
S. Lazzari, D. Mottet, and J.-L. Vercher, “Eye-hand coordination in rhythmical pointing,” Journal of motor behavior , vol. 41, no. 4, pp. 294–304, 2009
2009
Earlier work this paper cites.
M. C. Bowman, R. S. Johannson, and J. R. Flanagan, “Eye–hand coordination in a sequential target contact task,” Experimental brain research , vol. 195, no. 2, pp. 273–283, 2009
2009
Earlier work this paper cites.
E. D. Vidoni, J. S. McCarley, J. D. Edwards, and L. A. Boyd, “Manual and oculomotor performance develop contemporaneously but independently during continuous tracking,” Experimental brain research , vol. 195, no. 4, pp. 611–620, 2009
2009
Earlier work this paper cites.
T. Judd, K. Ehinger, F. Durand, and A. Torralba, “Learning to predict where humans look,” in IEEE International Conference on Computer Vision (ICCV) . IEEE, 2009, pp. 2106–2113
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
A. Borji, D. N. Sihite, and L. Itti, “Probabilistic learning of task-specific visual attention,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2012
2012
Earlier work this paper cites.
K. Yamada, Y. Sugano, T. Okabe, Y. Sato, A. Sugimoto, and K. Hiraki, “Attention prediction in egocentric video using motion and visual saliency,” in Advances in Image and Video Technology , Y.-S. Ho, Ed., 2012, pp. 277–288
2012
Earlier work this paper cites.
Y. J. Lee, J. Ghosh, and K. Grauman, “Discovering important people and objects for egocentric video summarization,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 2012, pp. 1346–1353
2012
Earlier work this paper cites.
H. Pirsiavash and D. Ramanan, “Detecting activities of daily living in first-person camera views.” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2012, pp. 2847–2854
2012
Earlier work this paper cites.
Y. Li, A. Fathi, and J. M. Rehg, “Learning to predict gaze in egocentric video,” in IEEE International Conference on Computer Vision (ICCV) , 2013
2013
Cited alongside, same era.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in European Conference on Computer Vision (ECCV) . Springer, 2014, pp. 740–755
2014
Cited alongside, same era.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3431–3440
2015
Cited alongside, same era.
X. Huang, C. Shen, X. Boix, and Q. Zhao, “Salicon: Reducing the semantic gap in saliency prediction by adapting deep neural networks,” in IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 262–270
2015
Cited alongside, same era.
M. Zhang, K. T. Ma, J. H. Lim, Q. Zhao, and J. Feng, “Deep future gaze: Gaze anticipation on egocentric videos using adversarial networks,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017
2017
Later among the works it cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-cnn,” in IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 2961–2969
2017
Later among the works it cites.
F. Chollet, J. Allaire, et al. , “R interface to keras,” https://github.com/rstudio/keras , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 1–9
2015
Cited alongside, same era.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in Advances in Neural Information Processing Systems (NeurIPS) , 2015, pp. 91–99
2015
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International Conference on Machine Learning (ICML) , 2015
2015
Cited alongside, same era.
N. Liu and J. Han, “Dhsnet: Deep hierarchical saliency network for salient object detection,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 678–686
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Defferrard, X. Bresson, and P. Vandergheynst, “Convolutional neural networks on graphs with fast localized spectral filtering,” in Advances in Neural Information Processing Systems (NeurIPS) , 2016, pp. 3844–3852
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Cited alongside, same era.
A. Palazzi, D. Abati, F. Solera, R. Cucchiara, et al. , “Predicting the driver’s focus of attention: the dr (eye) ve project,” IEEE Transactions on Pattern Analysis and Machine Intelligence (PAMI) , vol. 41, no. 7, pp. 1720–1733, 2018
2018
Later among the works it cites.
A. Tawari, P. Mallela, and S. Martin, “Learning to attend to salient targets in driving videos using fully convolutional rnn,” in International Conference on Intelligent Transportation Systems (ITSC) . IEEE, 2018, pp. 3225–3232
2018
Later among the works it cites.
Y. Xia, D. Zhang, J. Kim, K. Nakayama, K. Zipser, and D. Whitney, “Predicting driver attention in critical situations,” in Asian Conference on Computer Vision . Springer, 2018, pp. 658–674
2018
Later among the works it cites.
Z. Zhang, S. Bambach, C. Yu, and D. J. Crandall, “From coarse attention to fine-grained gaze: A two-stage 3d fully convolutional network for predicting eye gaze in first person video,” in British Machine Vision Conference (BMVC) , 2018
2018
Later among the works it cites.
Y. Huang, M. Cai, Z. Li, and Y. Sato, “Predicting gaze in egocentric video by learning task-dependent attention transition,” in European Conference on Computer Vision (ECCV) , 2018, pp. 754–769
2018
Later among the works it cites.
X. Wang and A. Gupta, “Videos as space-time region graphs,” in European Conference on Computer Vision (ECCV) , 2018, pp. 399–417
2018
Later among the works it cites.
Y. Li and A. Gupta, “Beyond grids: Learning graph representations for visual recognition,” in Advances in Neural Information Processing Systems (NeurIPS) , 2018, pp. 9225–9235
2018
Later among the works it cites.
J. Yang, J. Lu, S. Lee, D. Batra, and D. Parikh, “Graph r-cnn for scene graph generation,” in European Conference on Computer Vision (ECCV) , 2018, pp. 670–685
2018
Later among the works it cites.
R. Girshick, I. Radosavovic, G. Gkioxari, P. Dollár, and K. He, “Detectron,” https://github.com/facebookresearch/detectron , 2018
2018
Later among the works it cites.
2019
Later among the works it cites.
Z. Zhang, C. Yu, and D. Crandall, “A self validation network for object-level human attention estimation,” in Advances in Neural Information Processing Systems , 2019, pp. 14 702–14 713
2019
Later among the works it cites.
Y. Chen, M. Rohrbach, Z. Yan, Y. Shuicheng, J. Feng, and Y. Kalantidis, “Graph-based global reasoning networks,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 433–442
2019
Later among the works it cites.