Fetching the paper…
Reading the bibliography…
Speaker diarization(SD) is a classic task in speech processing and is crucial in multi-party scenarios such as meetings and conversations.
Efficient algorithms for agglomerative hierarchical clustering methods
William H. E. Day and Herbert Edelsbrunner. 1984 · 1984
Earlier work this paper cites.
The icsi meeting corpus
Adam L. Janin, Don Baron, Jane Edwards, Daniel P. W. Ellis, David Gelbart, Nelson Morgan, Barbara Peskin, Thilo Pfau, Elizabeth Shriberg, Andreas Stolcke, and Chuck Wooters. 2003 · 2003
Earlier work this paper cites.
Shota Horiguchi, Yusuke Fujita, Shinji Watanabe, Yawen Xue, and Kenji Nagamatsu. 2020 · 2005
Earlier work this paper cites.
The ami meeting corpus: A pre-announcement
Jean Carletta, Simone Ashby, Sebastien Bourban, Mike Flynn, Mael Guillemot, Thomas Hain, Jaroslav Kadlec, Vasilis Karaiskos, Wessel Kraaij, Melissa Kronenthal, et al. 2006 · 2006
Earlier work this paper cites.
Universal asr: Unifying streaming and non-streaming asr using a single encoder-decoder model
Zhifu Gao, Shiliang Zhang, Ming Lei, and Ian Mcloughlin. 2020 · 2010
Earlier work this paper cites.
Front-end factor analysis for speaker verification
Najim Dehak, Patrick Kenny, Réda Dehak, Pierre Dumouchel, and Pierre Ouellet. 2011 · 2011
Earlier work this paper cites.
Speaker role contextual modeling for language understanding and dialogue policy learning
Ta-Chung Chi, Po-Chun Chen, Shang-Yu Su, and Yun-Nung Chen. 2017a · 2017
Earlier work this paper cites.
Montreal forced aligner: Trainable text-speech alignment using kaldi
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner, and Morgan Sonderegger. 2017 · 2017
Earlier work this paper cites.
End-to-end text-independent speaker verification with triplet loss on short utterances
Chunlei Zhang and Kazuhito Koishida. 2017 · 2017
Earlier work this paper cites.
X-vectors: Robust dnn embeddings for speaker recognition
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur. 2018 · 2018
Earlier work this paper cites.
Speaker diarization with lstm
Quan Wang, Carlton Downey, Li Wan, P. A. Mansfield, and Ignacio Lopez-Moreno. 2017 · 2018
Cited alongside, same era.
Addressee and response selection in multi-party conversations with speaker interaction rnns
Rui Zhang, Honglak Lee, Lazaros Polymenakos, and Dragomir Radev. 2018 · 2018
Cited alongside, same era.
End-to-end neural speaker diarization with self-attention
Yusuke Fujita, Naoyuki Kanda, Shota Horiguchi, Yawen Xue, Kenji Nagamatsu, and Shinji Watanabe. 2019 · 2019
Cited alongside, same era.
Controllable time-delay transformer for real-time punctuation prediction and disfluency detection
Qian Chen, Mengzhe Chen, Bo Li, and Wen Wang. 2020 · 2020
Cited alongside, same era.
Ecapa-tdnn: Emphasized channel attention, propagation and aggregation in tdnn based speaker verification
Brecht Desplanques, Jenthe Thienpondt, and Kris Demuynck. 2020 · 2020
Cited alongside, same era.
A review of speaker diarization: Recent advances with deep learning
Tae Jin Park, Naoyuki Kanda, Dimitrios Dimitriadis, Kyu J. Han, Shinji Watanabe, and Shrikanth S. Narayanan. 2021 · 2021
Later among the works it cites.
Cam: Context-aware masking for robust speaker verification
Ya-Qi Yu, Siqi Zheng, Hongbin Suo, Yun Lei, and Wu-Jun Li. 2021 · 2021
Later among the works it cites.
A real-time speaker diarization system based on spatial spectrum
Siqi Zheng, Weilong Huang, Xianliang Wang, Hongbin Suo, Jinwei Feng, and Zhijie Yan. 2021 · 2021
Later among the works it cites.
Qmsum: A new benchmark for query-based multi-domain meeting summarization
Ming Zhong, Da Yin, Tao Yu, Ahmad Zaidi, Mutethia Mutuma, Rahul Jha, Ahmed Hassan Awadallah, Asli Celikyilmaz, Yang Liu, Xipeng Qiu, and Dragomir R. Radev. 2021 · 2021
Later among the works it cites.
Bertraffic: Bert-based joint speaker role and speaker change detection for air traffic control communications
Juan Zuluaga-Gomez, Seyyed Saeed Sarfjoo, Amrutha Prasad, Iuliia Nigmatulina, Petr Motlícek, Karel Ondrej, Oliver Ohneiser, and Hartmut Helmke. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Auto-tuning spectral clustering for speaker diarization using normalized maximum eigengap
Tae Jin Park, Kyu J. Han, Manoj Kumar, and Shrikanth S. Narayanan. 2020 · 2020
Cited alongside, same era.
Ecapa-tdnn embeddings for speaker diarization
Nauman Dawalatabad, Mirco Ravanelli, Franccois Grondin, Jenthe Thienpondt, Brecht Desplanques, and Hwidong Na. 2021 · 2021
Cited alongside, same era.
Yihui Fu, Luyao Cheng, Shubo Lv, Yukai Jv, Yuxiang Kong, Zhuo Chen, Yanxin Hu, Lei Xie, Jian Wu, Hui Bu, et al. 2021 · 2021
Cited alongside, same era.
Target-speaker voice activity detection with improved i-vector estimation for unknown number of speaker
Maokui He, Desh Raj, Zili Huang, Jun Du, Zhuo Chen, and Shinji Watanabe. 2021 · 2021
Cited alongside, same era.
Speaker role contextual modeling for language understanding and dialogue policy learning
Ta-Chung Chi, Po-Chun Chen, Shang-Yu Su, and Yun-Nung (Vivian) Chen. 2017b
Cited in the paper.
Zhihao Du, Shiliang Zhang, Siqi Zheng, and Zhijie Yan. 2022a
Cited in the paper.
Speaker overlap-aware neural diarization for multi-party meeting analysis
Zhihao Du, Shiliang Zhang, Siqi Zheng, and Zhijie Yan. 2022b
Cited in the paper.
Later among the works it cites.
Multimodal clustering with role induced constraints for speaker diarization
Nikolaos Flemotomos and Shrikanth S. Narayanan. 2022 · 2022
Later among the works it cites.
M2met: The icassp 2022 multi-channel multi-party meeting transcription challenge
Fan Yu, Shiliang Zhang, Yihui Fu, Lei Xie, Siqi Zheng, Zhihao Du, Weilong Huang, Pengcheng Guo, Zhijie Yan, Bin Ma, Xin Xu, and Hui Bu. 2022 · 2022
Later among the works it cites.
Reformulating speaker diarization as community detection with emphasis on topological structure
Siqi Zheng and Hongbin Suo. 2022 · 2022
Later among the works it cites.
PRISM: pre-trained indeterminate speaker representation model for speaker diarization and speaker verification
Siqi Zheng, Hongbin Suo, and Qian Chen. 2022 · 2022
Later among the works it cites.