Fetching the paper…
Reading the bibliography…
We propose a novel framework, called Disjoint Mapping Network (DIMNet), for cross-modal biometric matching, in particular of voices and faces.
Neuro-cognitive processing of faces and voices
A. W. Ellis · 1989
Earlier work this paper cites.
When eyewitnesses are also earwitnesses: Effects on visual and voice identifications
H. A. McAllister, R. H. Dale, N. J. Bregman, A. McCabe, and C. R. Cotton · 1993
Earlier work this paper cites.
Putting the face to the voice’: Matching identity across modality
M. Kamachi, H. Hill, K. Lander, and E. Vatikiotis-Bateson · 2003
Earlier work this paper cites.
An introduction to mds
F. Wickelmaier · 2003
Earlier work this paper cites.
Thinking the voice: neural correlates of voice perception
P. Belin, S. Fecteau, and C. Bedard · 2004
Earlier work this paper cites.
Hearing facial identities
S. R. Schweinberger, D. Robertson, and J. M. Kaufmann · 2007
Earlier work this paper cites.
Evaluation in information retrieval
C. D. Manning, P. Raghavan, and H. Schütze · 2008
Earlier work this paper cites.
Visualizing data using t-sne
L. van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
Understanding voice perception
P. Belin, P. Bestelmeyer, M. Latinus, and R. Watson · 2011
Cited alongside, same era.
The kaldi speech recognition toolkit
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, et al · 2011
Cited alongside, same era.
Hearing facial identities: Brain correlates of face–voice integration in person identification
S. R. Schweinberger, N. Kloth, and D. M. Robertson · 2011
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Deep face recognition
O. M. Parkhi, A. Vedaldi, A. Zisserman, et al · 2015
Joint face detection and alignment using multitask cascaded convolutional networks
K. Zhang, Z. Zhang, Z. Li, and Y. Qiao · 2016
Later among the works it cites.
Identification of individuals by trait prediction using whole-genome sequencing data
C. Lippert, R. Sabatini, M. C. Maher, E. Y. Kang, S. Lee, O. Arikan, A. Harley, A. Bernal, P. Garst, V. Lavrenko, et al · 2017
Later among the works it cites.
Sphereface: Deep hypersphere embedding for face recognition
W. Liu, Y. Wen, Z. Yu, M. Li, B. Raj, and L. Song · 2017
Later among the works it cites.
Voxceleb: a large-scale speaker identification dataset
A. Nagrani, J. S. Chung, and A. Zisserman · 2017
Later among the works it cites.
On learning associations of faces and voices
C. Kim, H. V. Shin, T.-H. Oh, K. Alexandre, M. Elgharib, and W. Matusik · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Large-margin softmax loss for convolutional neural networks
W. Liu, Y. Wen, Z. Yu, and M. Yang · 2016
Cited alongside, same era.
A discriminative feature learning approach for deep face recognition
Y. Wen, K. Zhang, Z. Li, and Y. Qiao · 2016
Cited alongside, same era.
Learnable pins: Cross-modal embeddings for person identity
A. Nagrani, S. Albanie, and A. Zisserman · 2018
Closest in time.
Seeing voices and hearing faces: Cross-modal biometric matching
A. Nagrani, S. Albanie, and A. Zisserman · 2018
Closest in time.
Optimal strategies for matching and retrieval problems by comparing covariates
Y. Wen, M. Al Ismail, B. Raj, and R. Singh · 2018
Closest in time.