Fetching the paper…
Reading the bibliography…
Human visual system can selectively attend to parts of a scene for quick perception, a biological mechanism known as Human attention.
C. W. Eriksen and J. E. Hoffman, “Temporal and spatial characteristics of selective encoding from visual displays,” Perception & Psychophysics , vol. 12, no. 2, pp. 201–204, 1972
1972
Earlier work this paper cites.
A. M. Treisman and G. Gelade, “A feature-integration theory of attention,” Cognitive Psychology , vol. 12, no. 1, pp. 97–136, 1980
1980
Earlier work this paper cites.
C. Koch and S. Ullman, “Shifts in selective visual attention: Towards the underlying neural circuitry,” in Matters of Intelligence , 1987, pp. 115–141
1987
Earlier work this paper cites.
J. M. Wolfe, K. R. Cave, and S. L. Franzel, “Guided search: An alternative to the feature integration model for visual search.” Journal of Experimental Psychology: Human Perception and Performance , vol. 15, no. 3, p. 419, 1989
1989
Earlier work this paper cites.
J. K. Tsotsos, S. M. Culhane, W. Y.K. Wai, Y. Lai, N. Davis, F. Nuflo, “Modeling visual attention via selective tuning,” in Artificial Intelligence , vol. 78, no. 1-2, pp. 507–545, 1995
1995
Earlier work this paper cites.
L. Itti, C. Koch, and E. Niebur, “A model of saliency-based visual attention for rapid scene analysis,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 20, no. 11, pp. 1254–1259, 1998
1998
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
L. Itti and C. Koch, “Computational modelling of visual attention,” Nature Reviews Neuroscience , vol. 2, no. 3, p. 194, 2001
2001
Earlier work this paper cites.
C. E. Connor, H. E. Egeth, and S. Yantis, “Visual attention: Bottom-up versus top-down,” Current Biology , vol. 14, no. 19, pp. 850–852, 2004
2004
Earlier work this paper cites.
D. Gao and N. Vasconcelos, “Discriminant saliency for visual recognition from cluttered scenes,” in Proc. Advances Neural Inf. Process. Syst. , 2005, pp. 481–488
2005
Earlier work this paper cites.
K. Koch, J. McLean, R. Segev, M. A. Freed, M. J. Berry II, V. Balasubramanian, and P. Sterling, “How much the eye tells the brain,” Current Biology , vol. 16, no. 14, pp. 1428–1434, 2006
2006
Earlier work this paper cites.
N. Bruce and J. Tsotsos, “Saliency based on information maximization,” in Proc. Advances Neural Inf. Process. Syst. , 2006, pp. 155–162
2006
Earlier work this paper cites.
C. Zach, T. Pock, and H. Bischof, “A duality based approach for realtime tv-l 1 optical flow,” in Joint Pattern Recognition Symposium , 2007, pp. 214–223
2007
Earlier work this paper cites.
M. D. Rodriguez, J. Ahmed, and M. Shah, “Action mach a spatio-temporal maximum average correlation height filter for action recognition,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2008, pp. 1–8
2008
Earlier work this paper cites.
A. D. Hwang, E. C. Higgins, and M. Pomplun, “A model of top-down attentional control during visual search in complex scenes,” Journal of Vision , vol. 9, no. 5, pp. 25–25, 2009
2009
Earlier work this paper cites.
T. Judd, K. Ehinger, F. Durand, and A. Torralba, “Learning to predict where humans look,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2009, pp. 2106–2113
2009
Earlier work this paper cites.
M. Marszalek, I. Laptev, and C. Schmid, “Actions in context,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2009, pp. 2929–2936
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “ImageNet: A Large-Scale Hierarchical Image Database,” in CVPR , 2009, pp. 248–255
2009
Earlier work this paper cites.
R. Achanta, S. Hemami, F. Estrada, and S. Susstrunk, “Frequency-tuned salient region detection,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2009, pp. 1597–1604
2009
Earlier work this paper cites.
M. Everingham, L. V. Gool, C. K. I. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (VOC) challenge,” Int. J. Comput. Vis. , vol. 88, no. 2, pp. 303–338, 2010
2010
Earlier work this paper cites.
P. Welinder, S. Branson, T. Mita, C. Wah, F. Schroff, S. Belongie, and P. Perona, “Caltech-UCSD Birds 200,” California Institute of Technology, Tech. Rep. CNS-TR-2010-001, 2010
2010
Earlier work this paper cites.
M. Carrasco, “Visual attention: The past 25 years,” Vision Research , vol. 51, no. 13, pp. 1484–1525, 2011
2011
Earlier work this paper cites.
P. K. Mital, T. J. Smith, R. L. Hill, and J. M. Henderson, “Clustering of gaze during dynamic scene viewing is predicted by motion,” Cognitive Computation , vol. 3, no. 1, pp. 5–24, 2011
2011
Earlier work this paper cites.
G. Sharma, F. Jurie, and C. Schmid, “Discriminative spatial saliency for image classification,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2012, pp. 3506–3513
2012
Earlier work this paper cites.
H. Hadizadeh, M. J. Enriquez, and I. V. Bajic, “Eye-tracking database for a set of standard video sequences,” IEEE Trans. Image Process. , vol. 21, no. 2, pp. 898–903, 2012
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Proc. Advances Neural Inf. Process. Syst. , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
T. Judd, F. Durand, and A. Torralba, “A benchmark of computational models of saliency to predict human fixations,” MIT Technical Report , 2012
2012
Earlier work this paper cites.
C. Yang, L. Zhang, H. Lu, X. Ruan, and M.-H. Yang, “Saliency detection via graph-based manifold ranking,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2013
2013
Earlier work this paper cites.
Y. Pinto, A. R. van der Leij, I. G. Sligte, V. A. Lamme, and H. S. Scholte, “Bottom-up and top-down attention are independent,” Journal of Vision , vol. 13, no. 3, pp. 16–16, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
A. Borji, H. R. Tavakoli, D. N. Sihite, and L. Itti, “Analysis of scores, datasets, and models in visual saliency prediction,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2013, pp. 921–928
2013
Earlier work this paper cites.
F. Katsuki and C. Constantinidis, “Bottom-up and top-down attention: Different processes and overlapping neural systems,” The Neuroscientist , vol. 20, no. 5, pp. 509–521, 2014
2014
Cited alongside, same era.
Y. Li, X. Hou, C. Koch, J. M. Rehg, and A. L. Yuille, “The secrets of salient object segmentation,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2014, pp. 280–287
2014
Cited alongside, same era.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in Proc. Advances Neural Inf. Process. Syst. , 2014, pp. 568–576
2014
Cited alongside, same era.
V. Mnih, N. Heess, A. Graves, and K. Kavukcuoglu, “Recurrent models of visual attention,” in Proc. Advances Neural Inf. Process. Syst. , 2014, pp. 2204–2212
2014
Cited alongside, same era.
N. Karessli, Z. Akata, B. Schiele, and A. Bulling, “Gaze embeddings for zero-shot image classification,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2017
2017
Later among the works it cites.
W. Wang, and J. Shen, “Deep visual attention prediction,” IEEE Trans. Image Process. , vol. 27, no. 5, pp. 2368–2378, 2017
2017
Later among the works it cites.
S. Zagoruyko and N. Komodakis, “Paying more attention to attention: Improving the performance of convolutional neural networks via attention transfer,” in Proc. Int. Conf. Learn. Representations , 2017
2017
Later among the works it cites.
S. Song, C. Lan, J. Xing, W. Zeng, and J. Liu, “An end-to-end spatio-temporal attention model for human action recognition from skeleton data,” in AAAI Conference on Artificial Intelligence , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
A. Borji, M.-M. Cheng, H. Jiang, and J. Li, “Salient object detection: A benchmark,” IEEE Trans. Image Process. , vol. 24, no. 12, pp. 5706–5722, 2015
2015
Cited alongside, same era.
S. Mathe and C. Sminchisescu, “Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 37, no. 7, pp. 1408–1424, 2015
2015
Cited alongside, same era.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in Proc. Int. Conf. Learn. Representations , 2015
2015
Cited alongside, same era.
A. M. Rush, S. Chopra, and J. Weston, “A neural attention model for abstractive sentence summarization,” in Proceedings of Conference on Empirical Methods in Natural Language Processing , 2015, pp. 379–389
2015
Cited alongside, same era.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in Proc. Int. Conf. Learn. Representations , 2015
2015
Cited alongside, same era.
C. Cao, X. Liu, Y. Yang, Y. Yu, J. Wang, Z. Wang, Y. Huang, L. Wang, C. Huang, W. Xu, D. Ramanan, and T. S. Huang, “Look and think twice: Capturing top-down visual attention with feedback convolutional neural networks,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2015
2015
Cited alongside, same era.
S. Sharma, R. Kiros, and R. Salakhutdinov, “Action recognition using visual attention,” in Proc. Int. Conf. Learn. Representations Workshops. , 2015
2015
Cited alongside, same era.
2017
Later among the works it cites.
2017
Later among the works it cites.
W. Wang, Y. Xu, J. Shen, and S. Zhu, “Attentive fashion grammar network for fashion landmark detection and clothing category classification,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , pp. 4271–4280, 2018
2018
Later among the works it cites.
W. Wang, J. Shen, R. Yang, and F. Porikli, “Saliency-aware video object segmentation,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 40, no. 1, pp. 20–33, 2018
2018
Later among the works it cites.
J. Zhang, S. A. Bargal, Z. Lin, J. Brandt, X. Shen, S. Sclaroff, Stan, “Top-down neural attention by excitation backprop,” in Int. J. Comput. Vis. , vol. 126, no. 10, pp. 1084–1102, 2018
2018
Later among the works it cites.
M. M. Farazi and S. Khan, “Reciprocal attention fusion for visual question answering,” British Machine Vision Conference , 2018
2018
Later among the works it cites.
C. Zhu, X. Tan, F. Zhou, X. Liu, K. Yue, E. Ding, and Y. Ma, “Fine-grained video categorization with redundancy reduction attention,” in Proc. Eur. Conf. Comput. Vis. , 2018, pp. 136–152
2018
Later among the works it cites.
Y. Du, C. Yuan, B. Li, L. Zhao, Y. Li, and W. Hu, “Interaction-aware spatio-temporal pyramid attention networks for action classification,” in Proc. Eur. Conf. Comput. Vis. , 2018, pp. 373–389
2018
Later among the works it cites.
X. Zhang, T. Wang, J. Qi, H. Lu, and G. Wang, “Progressive attention guided recurrent network for salient object detection,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2018
2018
Later among the works it cites.
W. Wang, J. Shen, X. Dong, and A. Borji, “Salient object detection driven by fixation prediction,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2018, pp. 1711–1720
2018
Later among the works it cites.
N. Liu, J. Han, and M.-H. Yang, “Picanet: Learning pixel-wise contextual attention for saliency detection,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2018
2018
Later among the works it cites.
S. Chen, X. Tan, B. Wang, and X. Hu, “Reverse attention for salient object detection,” in Proc. Eur. Conf. Comput. Vis. , 2018
2018
Later among the works it cites.
S. Woo, J. Park, J.-Y. Lee, and I. So Kweon, “CBAM: Convolutional block attention module,” in Proc. Eur. Conf. Comput. Vis. , 2018
2018
Later among the works it cites.
Z. C. Lipton, “The mythos of model interpretability,” Queue , vol. 16, no. 3, pp. 30:31–30:57, 2018
2018
Later among the works it cites.
D. Xu, W. Wang, H. Tang, H. Liu, N. Sebe, and E. Ricci“Structured attention guided convolutional neural fields for monocular depth estimation,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2018, pp. 3917–3925
2018
Later among the works it cites.
W. Wang, J. Shen, F. Guo, M.-M. Cheng, and A. Borji, “Revisiting video saliency: A large-scale benchmark and a new model,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2018
2018
Later among the works it cites.
S. Jetley, N. A. Lord, N. Lee, and P. H. Torr, “Learn to pay attention,” in Proc. Int. Conf. Learn. Representations , 2018
2018
Later among the works it cites.
P. Zhang, J. Xue, C. Lan, W. Zeng, Z. Gao, and N. Zheng, “Adding attentiveness to the neurons in recurrent neural networks,” in Proc. Eur. Conf. Comput. Vis. , 2018
2018
Later among the works it cites.
Z. Bylinskii, T. Judd, A. Oliva, A. Torralba, and F. Durand, “What do different evaluation metrics tell us about saliency models?” IEEE Trans. Pattern Anal. Mach. Intell. , 2018
2018
Later among the works it cites.
W. Wang, S. Zhao, J. Shen, S. C. Hoi, and A. Borji, “Salient object detection with pyramid attention and salient edges,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , pp. 1448–1457, 2019
2019
Closest in time.
W. Wang, H. Song, S. Zhao, J. Shen, S. Zhao, S. C. Hoi, and H. Ling, “Learning unsupervised video object segmentation through visual attention,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , pp. 3064–3074, 2019
2019
Closest in time.
C. Xie, S. Liu, C. Li, M.-M. Cheng, W. Zuo, X. Liu, S. Wen, and E. Ding, “Image Inpainting with Learnable Bidirectional Attention Maps”, in Proc. IEEE Int. Conf. Comput. Vis. , 2019, pp. 8858–8867
2019
Closest in time.
M. Zhai, X. Xiang, R. Zhang, N. Lv, and A. El Saddik, “Optical Flow Estimation Using Dual Self-Attention Pyramid Networks,” IEEE Trans. Circuits Syst. Video Technol. , 2019
2019
Closest in time.
2019
Closest in time.
W. Wang, X. Lu, J. Shen, D. Crandall, and L. Shao, “Zero-shot video object segmentation via attentive graph neural networks,” in Proc. IEEE Int. Conf. Comput. Vis. , pp. 9236–9245, 2019
2019
Closest in time.
V. Navalpakkam and L. Itti, “An integrated model of top-down and bottom-up attention for optimizing detection speed,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2006, pp. 2049–2056
2056
Closest in time.