H. McGurk and J. MacDonald, “Hearing lips and seeing voices,” Nat. , vol. 264, no. 5588, p. 746, 1976
1976
Earlier work this paper cites.
H. P. Moravec, “Sensor fusion in certainty grids for mobile robots,” in Sensor devices and systems for robotics . Springer, 1989, pp. 253–276
1989
Earlier work this paper cites.
J. R. Hershey and J. R. Movellan, “Audio vision: Using audio-visual synchrony to locate sounds,” in NeurIPS , 2000
2000
Earlier work this paper cites.
J. W. Fisher III, T. Darrell, W. T. Freeman, and P. A. Viola, “Learning joint statistical models for audio-visual fusion and segregation,” in NeurIPS , 2001
2001
Earlier work this paper cites.
H. L. Van Trees, Optimum array processing: Part IV of detection, estimation, and modulation theory . John Wiley & Sons, 2004
2004
Earlier work this paper cites.
S. Thrun, W. Burgard, and D. Fox, Probabilistic robotics . MIT press, 2005
2005
Earlier work this paper cites.
E. Kidron, Y. Y. Schechner, and M. Elad, “Pixels that sound,” in CVPR , 2005
2005
Earlier work this paper cites.
K. P. Körding, U. Beierholm, W. J. Ma, S. Quartz, J. B. Tenenbaum, and L. Shams, “Causal inference in multisensory perception,” PLOS ONE , vol. 2, no. 9, p. e943, 2007
2007
Earlier work this paper cites.
Z. Barzelay and Y. Y. Schechner, “Harmony in motion,” in CVPR , 2007
2007
Earlier work this paper cites.
H. Izadinia, I. Saleemi, and M. Shah, “Multimodal analysis for identification and segmentation of moving-sounding objects,” IEEE TMM , vol. 15, no. 2, pp. 378–390, 2013
2013
Earlier work this paper cites.
A. Zunino, M. Crocco, S. Martelli, A. Trucco, A. Del Bue, and V. Murino, “Seeing the sound: A new multimodal imaging device for computer vision,” in CVPR Workshop , 2015
2015
Earlier work this paper cites.
H. Mei, M. Bansal, and M. R. Walter, “Listen, attend, and walk: Neural mapping of navigational instructions to action sequences.” in AAAI , 2016
2016
Earlier work this paper cites.
A. Miller, A. Fisch, J. Dodge, A.-H. Karimi, A. Bordes, and J. Weston, “Key-value memory networks for directly reading documents,” in EMNLP , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in ICML , 2016
2016
Earlier work this paper cites.