Fetching the paper…
Reading the bibliography…
Disentangling uncorrelated information in speech utterances is a crucial research topic within speech community.
Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J.G. Fiscus, and D. S. Pallett · 1993
Earlier work this paper cites.
Technical Report , 2012
The nist year 2012 speaker recognition evaluation plan · 2012
Earlier work this paper cites.
Librispeech: An ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Earlier work this paper cites.
The speakers in the wild (SITW) speaker recognition database
Mitchell McLaren, Luciana Ferrer, Diego Castán, and Aaron Lawson · 2016
Earlier work this paper cites.
Unsupervised learning of disentangled and interpretable representations from sequential data
Wei-Ning Hsu, Yu Zhang, and James R. Glass · 2017
Earlier work this paper cites.
Improving the effectiveness of speaker verification domain adaptation with inadequate in-domain data
Bengt J. Borgström, Elliot Singer, Douglas A. Reynolds, and Seyed Omid Sadjadi · 2017
Earlier work this paper cites.
Voxceleb: A large-scale speaker identification dataset
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman · 2017
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman · 2018
Earlier work this paper cites.
Learning to generalize: Meta-learning for domain generalization
Da Li, Yongxin Yang, Yi-Zhe Song, and Timothy M. Hospedales · 2018
Earlier work this paper cites.
Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factorization
Wei-Ning Hsu, Yu Zhang, Ron J. Weiss, Yu-An Chung, Yuxuan Wang, Yonghui Wu, and James R. Glass · 2019
Earlier work this paper cites.
Autoencoder-based semi-supervised curriculum learning for out-of-domain speaker verification
Siqi Zheng, Gang Liu, Hongbin Suo, and Yun Lei · 2019
Cited alongside, same era.
Adversarial domain adaptation for speaker verification using partially shared network
Zhengyang Chen, Shuai Wang, and Yanmin Qian · 2020
Cited alongside, same era.
Phonetically-aware coupled network for short duration text-independent speaker verification
Siqi Zheng, Yun Lei, and Hongbin Suo · 2020
Cited alongside, same era.
In-domain and out-of-domain data augmentation to improve children’s speaker verification system in limited data scenario
S. Shahnawazuddin, Waquar Ahmad, Nagaraj Adiga, and Avinash Kumar · 2020
Cited alongside, same era.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli · 2020
Cited alongside, same era.
AISHELL-4: an open source dataset for speech enhancement, separation, recognition and speaker diarization in conference scenario
Yihui Fu, Luyao Cheng, Shubo Lv, Yukai Jv, Yuxiang Kong, Zhuo Chen, Yanxin Hu, Lei Xie, Jian Wu, Hui Bu, Xin Xu, Jun Du, and Jingdong Chen · 2021
Later among the works it cites.
A real-time speaker diarization system based on spatial spectrum
Siqi Zheng, Weilong Huang, Xianliang Wang, Hongbin Suo, Jinwei Feng, and Zhijie Yan · 2021
Later among the works it cites.
Investigation of spatial-acoustic features for overlapping speech detection in multiparty meetings
Shiliang Zhang, Siqi Zheng, Weilong Huang, Ming Lei, Hongbin Suo, Jinwei Feng, and Zhijie Yan · 2021
Later among the works it cites.
Deep representation decomposition for rate-invariant speaker verification
Fuchuan Tong, Siqi Zheng, Haodong Zhou, Xingjia Xie, Qingyang Hong, and Lin Li · 2022
Later among the works it cites.
Investigating effective domain adaptation method for speaker verification task
Guangxing Li, Wangjin Zhou, Sheng Li, Yi Zhao, Jichen Yang, and Hao Huang · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cn-celeb: A challenging chinese speaker recognition dataset
Y. Fan, J. W. Kang, L. T. Li, K. C. Li, H. L. Chen, S. T. Cheng, P. Y. Zhang, Z. Y. Zhou, Y. Q. Cai, and D. Wang · 2020
Cited alongside, same era.
Differential beamforming for uniform circular array with directional microphones
Weilong Huang and Jinwei Feng · 2020
Cited alongside, same era.
ECAPA-TDNN: emphasized channel attention, propagation and aggregation in TDNN based speaker verification
Brecht Desplanques, Jenthe Thienpondt, and Kris Demuynck · 2020
Cited alongside, same era.
Deep feature cyclegans: Speaker identity preserving non-parallel microphone-telephone domain adaptation for speaker verification
Saurabh Kataria, Jesús Villalba, Piotr Zelasko, Laureano Moro-Velázquez, and Najim Dehak · 2021
Cited alongside, same era.
Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai, Kushal Lakhotia, Ruslan Salakhutdinov, and Abdelrahman Mohamed · 2021
Cited alongside, same era.
Wavlm: Large-scale self-supervised pre-training for full stack speech processing
Sanyuan Chen, Chengyi Wang, Zhengyang Chen, Yu Wu, Shujie Liu, Zhuo Chen, Jinyu Li, Naoyuki Kanda, Takuya Yoshioka, Xiong Xiao, Jian Wu, Long Zhou, Shuo Ren, Yanmin Qian, Yao Qian, Jian Wu, Michael Zeng, Xiangzhan Yu, and Furu Wei
Cited in the paper.
Beamtransformer: Microphone array-based overlapping speech detection
Siqi Zheng, Shiliang Zhang, Weilong Huang, Qian Chen, Hongbin Suo, Ming Lei, Jinwei Feng, and Zhijie Yan
Cited in the paper.
Learning domain-invariant transformation for speaker verification
Hanyi Zhang, Longbiao Wang, Kong Aik Lee, Meng Liu, Jianwu Dang, and Hui Chen · 2022
Later among the works it cites.
M2met: The icassp 2022 multi-channel multi-party meeting transcription challenge
Fan Yu, Shiliang Zhang, Yihui Fu, Lei Xie, Siqi Zheng, Zhihao Du, Weilong Huang, Pengcheng Guo, Zhijie Yan, Bin Ma, Xin Xu, and Hui Bu · 2022
Later among the works it cites.
CAM++: A fast and efficient network for speaker verification using context-aware masking
Hui Wang, Siqi Zheng, Yafeng Chen, Luyao Cheng, and Qian Chen · 2023
Closest in time.
An enhanced res2net with local and global feature fusion for speaker verification
Yafeng Chen, Siqi Zheng, Hui Wang, Luyao Cheng, Qian Chen, and Jiajun Qi · 2023
Closest in time.