Fetching the paper…
Reading the bibliography…
Recently, we proposed a novel speaker diarization method called End-to-End-Neural-Diarization-vector clustering (EEND-vector clustering) that integrates clustering-based and end-to-end neural network-based diarization approaches into one framework.
M. Przybocki and A. Martin, 2000 NIST Speaker Recognition Evaluation (LDC2001S97) . Philadelphia, New Jersey: Linguistic Data Consortium, 2001
2001
Earlier work this paper cites.
K. Wagstaff, C. Cardie, S. Rogers, and S. S. Schroedl, “Constrained k-means clustering with background knowledge,” in Proc. 18th International Conference on Machine Learning (ICML) , 2001
2001
Earlier work this paper cites.
S. D. Kamvar, D. Klein, and C. D. Manning, “Spectral leawrning,” in Proc. Proceedings ofthe 18th International Joint Conference on Artificial Intelligence (IJCAI) , 2003, pp. 561–566
2003
Earlier work this paper cites.
2005
Earlier work this paper cites.
J. Carletta, S. Ashby, S. Bourban, M. Flynn, M. Guillemot, T. Hain, J. Kadlec, V. Karaiskos, W. Kraaij, M. Kronenthal, G. Lathoud, M. Lincoln, A. Lisowska, I. McCowan, W. Post, D. Reidsma, , and P. Wellner, “The AMI meeting corpus: A pre-announcement,” in The Second International Conference on Machine Learning for Multimodal Interaction, ser. MLMI’05 , 2006, pp. 28–39
2006
Earlier work this paper cites.
U. von Luxburg, “A tutorial on spectral clustering,” Statist. and Comput. , vol. 17, no. 4, pp. 395–416, 2007
2007
Earlier work this paper cites.
2009
Earlier work this paper cites.
I. Davidson and S. S. Ravi, “Using instance-level constraints in agglomerative hierarchical clustering: theoretical and empirical results,” Data Mining and Knowledge Discovery , vol. 77, no. 18, pp. 257–282, Dec. 2009
2009
Earlier work this paper cites.
X. Anguera, S. Bozonnet, N. Evans, C. Fredouille, G. Friedland, and O. Vinyals, “Speaker diarization: A review of recent research,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 20, no. 2, pp. 356–370, Feb 2012
2012
Cited alongside, same era.
X. Wang, B. Qian, and I. Davidson, “On constrained spectral clustering and its applications,” Data Mining and Knowledge Discovery , vol. 28, pp. 1–30, Dec. 2014
2014
Cited alongside, same era.
2015
Cited alongside, same era.
D. Snyder, P. Ghahremani, D. Povey, D. Garcia-Romero, Y. Carmiel, , and S. Khudanpur, “Deep neural network-based speaker embeddings for end-to-end speaker verification,” in Proc. IEEE Spoken Language Technology Workshop , 2016
2016
Cited alongside, same era.
H. Aronowitz, W. Zhu, M. Suzuki, G. Kurata, and R. Hoory, “New advances in speaker diarization,” in Proc. Interspeech 2020 , 2018, pp. 279–283
2018
Later among the works it cites.
Y. Fujita, N. Kanda, S. Horiguchi, K. Nagamatsu, and S. Watanabe, “End-to-end neural speaker diarization with permutation-free objectives,” in Proc. Interspeech 2019 , 2019, pp. 4300–4304
2019
Later among the works it cites.
Y. Fujita, N. Kanda, S. Horiguchi, Y. Xue, K. Nagamatsu, and S. Watanabe, “End-to-end neural speaker diarization with self-attention,” in Proc. IEEE ASRU , 2019, pp. 296–303
2019
Later among the works it cites.
A. Zhang, Q. Wang, Z. Zhu, J. Paisley, and C. Wang, “Fully supervised speaker diarization,” in Proc. 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2019, pp. 6301–6305
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Ko, V. Peddinti, D. Povey, M. L. Seltzer, and S. Khudanpur, “A study on data augmentation of reverberant speech for robust speech recognition,” in Proc. 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , March 2017, pp. 5220––5224
2017
Cited alongside, same era.
N. Ryant, K. Church, C. Cieri, A. Cristia, J. Du, S. Ganapathy, and M. Liberman, First DIHARD Challenge Evaluation Plan , 2018, https://zenodo.org/record/1199638
2018
Cited alongside, same era.
G. Sell, D. Snyder, A. McCree, D. Garcia-Romero, J. Villalba, M. Maciejewski, V. Manohar, N. Dehak, D. Povey, S. Watanabe, and S. Khudanpur, “Diarization is hard: Some experiences and lessons learned for the JHU team in the inaugural DIHARD challenge,” in Proc. Interspeech 2018 , 2018, pp. 2808–2812. [Online]. Available: http://dx.doi.org/10.21437/Interspeech.2018-1893
2018
Cited alongside, same era.
M. Diez, F. Landini, L. Burget, J. Rohdin, A. Silnova, K. Zmolikova, O. Novotný, K. Veselý, O. Glembek, O. Plchot, L. Mošner, and P. Matějka, “BUT system for DIHARD speech diarization challenge 2018,” in Proc. Interspeech 2018 , 2018, pp. 2798–2802. [Online]. Available: http://dx.doi.org/10.21437/Interspeech.2018-1749
2018
Cited alongside, same era.
https://github.com/nttcslab-sp/EEND-vector-clustering
Cited in the paper.
Z. Huang, S. Watanabe, Y. Fujita, P. García, Y. Shao, D. Povey, and S. Khudanpur, “Speaker diarization with region proposal network,” in ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2020, pp. 6514–6518
2020
Later among the works it cites.
N. Ryant, P. Singh, V. Krishnamohan, R. Varma, K. Church, C. Cieri, J. Du, S. Ganapathy, and M. Liberman, “The third DIHARD diarization challenge,” 2021
2021
Closest in time.
S. Horiguchi, N. Yalta, P. Garcia, Y. Takashima, Y. Xue, D. Raj, Z. Huang, Y. Fujita, S. Watanabe, and S. Khudanpur, “The hitachi-jhu DIHARD III system: Competitive end-to-end neural diarization and x-vector clustering systems combined by DOVER-Lap,” 2021
2021
Closest in time.
K. Kinoshita, M. Delcroix, and N. Tawara, “Integrating end-to-end neural and clustering-based diarization: Getting the best of both worlds,” in Proc. 2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) (To appear) , 2021
2021
Closest in time.