Fetching the paper…
Reading the bibliography…
This paper studies audio-visual deep saliency prediction.
E. C. Cherry, “Some experiments on the recognition of speech, with one and with two ears,” The Journal of the Acoustical Society of America , vol. 25, no. 5, pp. 975–979, 1953
1953
Earlier work this paper cites.
E. E. Maccoby and K. W. Konrad, “Age trends in selective listening,” Journal of Experimental Child Psychology , vol. 3, no. 2, pp. 113 – 122, 1966
1966
Earlier work this paper cites.
A. M. Treisman and G. Gelade, “A feature-integration theory of attention,” Cognitive Psychology , vol. 12, no. 1, pp. 97 – 136, 1980
1980
Earlier work this paper cites.
D. J. Lewkowicz, “Sensory dominance in infants: Ii. ten-month-old infants’ response to auditory-visual compounds,” Developmental Psychology , vol. 24, no. 2, pp. 172–182, 1998
1998
Earlier work this paper cites.
J. M. Buhmann, J. Malik, and P. Perona, “Image recognition: Visual grouping, recognition, and learning,” Proceedings of the National Academy of Sciences , vol. 96, no. 25, pp. 14 203–14 204, 1999
1999
Earlier work this paper cites.
L. Itti and C. Koch, “A saliency-based search mechanism for overt and covert shifts of visual attention,” Vision Research , vol. 40, no. 10, pp. 1489 – 1506, 2000
2000
Earlier work this paper cites.
J. Richards, “Development of multimodal attention in young infants: modification of the startle reflex by attention,” Psychophysiology , vol. 37, no. 1, pp. 65–75, 2000
2000
Earlier work this paper cites.
J. Bartgis, A. R. Lilly, and D. G. Thomas, “Event-related potential and behavioral measures of attention in 5-, 7-, and 9-year-olds,” The Journal of General Psychology , vol. 130, no. 3, pp. 311–335, 2003
2003
Earlier work this paper cites.
S. N. Wrigley and G. J. Brown, “A computational model of auditory selective attention,” IEEE Transactions on Neural Networks , vol. 15, no. 5, pp. 1151–1163, 2004
2004
Earlier work this paper cites.
L. Itti, “Automatic foveation for video compression using a neurobiological model of visual attention,” IEEE Transactions on Image Processing , vol. 13, no. 10, pp. 1304–1318, 2004
2004
Earlier work this paper cites.
C. Kayser, C. I. Petkov, M. Lippert, and N. K. Logothetis, “Mechanisms for allocating auditory attention: An auditory saliency map,” Current Biology , vol. 15, no. 21, pp. 1943 – 1947, 2005
2005
Earlier work this paper cites.
S. Marat, M. Guironnet, and D. Pellerin, “Video summarization using a visual attention model,” in European Signal Processing Conference , 2007
2007
Earlier work this paper cites.
B. W. Tatler, “The central fixation bias in scene viewing: selecting an optimal viewing position independently of motor biases and image feature distributions.” Journal of vision , vol. 7 14, pp. 4.1–17, 2007
2007
Earlier work this paper cites.
E. Van der Burg, C. N. L. Olivers, A. W. Bronkhorst, and J. Theeuwes, “Audiovisual events capture attention: Evidence from temporal order judgments,” Journal of Vision , vol. 8, no. 5, pp. 2–2, 2008
2008
Earlier work this paper cites.
Yu Fu, Jian Cheng, Zhenglong Li, and Hanqing Lu, “Saliency cuts: An automatic approach to object segmentation,” in International Conference on Pattern Recognition (ICPR) , 2008
2008
Earlier work this paper cites.
L. Zhang, M. H. Tong, T. K. Marks, H. Shan, and G. W. Cottrell, “SUN: A Bayesian framework for saliency using natural statistics,” Journal of Vision , vol. 8, no. 7, pp. 32–32, 2008
2008
Earlier work this paper cites.
A. Mishra, Y. Aloimonos, and Cheong Loong Fah, “Active segmentation with fixation,” in International Conference on Computer Vision (ICCV) , 2009
2009
Earlier work this paper cites.
H. J. Seo and P. Milanfar, “Static and space-time visual saliency detection by self-resemblance,” Journal of Vision , vol. 9, no. 12, pp. 15–15, 2009
2009
Earlier work this paper cites.
C. Sandor, A. Cunningham, A. Dey, and V. Mattila, “An augmented reality x-ray system based on visual saliency,” in IEEE International Symposium on Mixed and Augmented Reality , 2010
2010
Earlier work this paper cites.
J. K. Tsotsos, A Computational Perspective on Visual Attention . The MIT Press, 2011
2011
Earlier work this paper cites.
M. Carrasco, “Visual attention: the past 25 years,” Vision research , vol. 41, no. 13, pp. 1484–1525, 2011
2011
Earlier work this paper cites.
P. K. Mital, T. J. Smith, R. L. Hill, and J. M. Henderson, “Clustering of gaze during dynamic scene viewing is predicted by motion,” Cognitive Computation , vol. 3, no. 1, pp. 5–24, 2011
2011
Earlier work this paper cites.
Q. Zhao and C. Koch, “Learning a saliency map using fixated locations in natural scenes,” Journal of Vision , vol. 11, no. 3, pp. 9–9, 2011
2011
Cited alongside, same era.
N. Mesgarani and E. Chang, “Selective cortical representation of attended speaker in multi-talker speech perception,” Nature , vol. 485, no. 7397, pp. 233–242, 2012
2012
Cited alongside, same era.
P. Bertolino, “Sensarea: An authoring tool to create accurate clickable videos,” in International Workshop on Content-Based Multimedia Indexing (CBMI) , 2012
2012
Cited alongside, same era.
T. Judd, F. Durand, and A. Torralba, “A benchmark of computational models of saliency to predict human fixations,” in MIT Technical Report , 2012
2012
Cited alongside, same era.
A. Borji and L. Itti, “State-of-the-art in visual attention modeling,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 35, no. 1, pp. 185–207, 2013
2016
Later among the works it cites.
2016
Later among the works it cites.
E. M. Kaya and M. Elhilali, “Modelling auditory attention,” Philosophical Transactions of the Royal Society of London B: Biological Sciences , vol. 372, no. 1714, 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
A. Borji, D. N. Sihite, and L. Itti, “What stands out in a scene? a study of human explicit saliency judgment,” Vision research , vol. 91, pp. 62–77, 2013
2013
Cited alongside, same era.
D. Oldoni, B. De Coensel, M. Boes, M. Rademaker, B. De Baets, T. Van Renterghem, and D. Botteldooren, “A computational model of auditory attention for use in soundscape research,” The Journal of the Acoustical Society of America , vol. 134, no. 1, pp. 852–861, 2013
2013
Cited alongside, same era.
A. Borji, H. R. Tavakoli, D. N. Sihite, and L. Itti, “Analysis of scores, datasets, and models in visual saliency prediction,” in IEEE International Conference on Computer Vision (ICCV) , 2013
2013
Cited alongside, same era.
H. Rezazadegan Tavakoli, E. Rahtu, and J. Heikkilä, “Spherical center-surround for video saliency detection using sparse sampling,” in Advanced Concepts for Intelligent Vision Systems , 2013
2013
Cited alongside, same era.
H. Hadizadeh and I. V. Bajić, “Saliency-aware video compression,” IEEE Transactions on Image Processing , vol. 23, no. 1, pp. 19–33, 2014
2014
Cited alongside, same era.
Y. Li, X. Hou, C. Koch, J. M. Rehg, and A. L. Yuille, “The secrets of salient object segmentation,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014
2014
Cited alongside, same era.
H. Rezazadegan Tavakoli, “Visual saliency and eye movement: modeling and applications,” Ph.D. dissertation, University of Oulu, 2014
2014
Cited alongside, same era.
2017
Later among the works it cites.
S. Hershey, S. Chaudhuri, D. P. W. Ellis, J. F. Gemmeke, A. Jansen, C. Moore, M. Plakal, D. Platt, R. A. Saurous, B. Seybold, M. Slaney, R. Weiss, and K. Wilson, “CNN architectures for large-scale audio classification,” in International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2017
2017
Later among the works it cites.
H. R. Tavakoli, F. Ahmed, A. Borji, and J. Laaksonen, “Saliency revisited: Analysis of mouse movements versus fixations,” in The IEEE Conference on Computer Vision and Pattern Recognition, (CVPR) , 2017
2017
Later among the works it cites.
J. Wang, H. R. Tavakoli, and J. Laaksonen, “Fixation prediction in videos using unsupervised hierarchical features,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , 2017
2017
Later among the works it cites.
V. Leborán Alvarez, A. García-Díaz, X. R. Fdez-Vidal, and X. M. Pardo, “Dynamic whitening saliency,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 5, pp. 893–907, 2017
2017
Later among the works it cites.
Z. Zhang, Y. Xu, J. Yu, and S. Gao, “Saliency detection in 360 ∘ 360^{\circ} videos,” in European Conference on Computer Vision (ECCV) , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
M. Amirul Islam, M. Kalash, and N. D. B. Bruce, “Revisiting salient object detection: Simultaneous detection, ranking, and subitizing of multiple salient objects,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Later among the works it cites.
W. Wang, J. Shen, F. Guo, M.-M. Cheng, and A. Borji, “Revisiting video saliency: A large-scale benchmark and a new model,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Later among the works it cites.
M. Cornia, L. Baraldi, G. Serra, and R. Cucchiara, “Predicting Human Eye Fixations via an LSTM-based Saliency Attentive Model,” IEEE Transactions on Image Processing , vol. 27, no. 10, pp. 5142–5154, 2018
2018
Later among the works it cites.
G. Boccignone, V. Cuculo, A. D’Amelio, G. Grossi, and R. Lanzarotti, “Give ear to my face: Modelling multimodal attention to social interactions,” in European Conference on Computer Vision Workshops , 2018
2018
Later among the works it cites.
L. Jiang, M. Xu, T. Liu, M. Qiao, and Z. Wang, “DeepVS: A deep learning based video saliency prediction approach,” in European Conference on Computer Vision (ECCV) , 2018
2018
Later among the works it cites.
K. Hara, H. Kataoka, and Y. Satoh, “Can spatiotemporal 3d cnns retrace the history of 2d cnns and imagenet?” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
M. Kümmerer, T. S. A. Wallis, and M. Bethge, “Saliency benchmarking made easy: Separating models, maps and metrics,” in European Conference on Computer Vision (ECCV) , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
D.-P. Fan, W. Wang, M.-M. Cheng, and J. Shen, “Shifting more attention to video salient object detection,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Closest in time.