Fetching the paper…
Reading the bibliography…
The training of modern speech processing systems often requires a large amount of simulated room impulse response (RIR) data in order to allow the systems to generalize well in real-world, reverberant environments.
J. B. Allen and D. A. Berkley, “Image method for efficiently simulating small-room acoustics,” The Journal of the Acoustical Society of America , vol. 65, no. 4, pp. 943–950, 1979
1979
Earlier work this paper cites.
A. Kulowski, “Algorithmic representation of the ray tracing technique,” Applied Acoustics , vol. 18, no. 6, pp. 449–469, 1985
1985
Earlier work this paper cites.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (PESQ) - a new method for speech quality assessment of telephone networks and codecs,” in Acoustics, Speech and Signal Processing (ICASSP), 2001 IEEE International Conference on , vol. 2. IEEE, 2001, pp. 749–752
2001
Earlier work this paper cites.
T. Funkhouser, N. Tsingos, I. Carlbom, G. Elko, M. Sondhi, J. E. West, G. Pingali, P. Min, and A. Ngan, “A beam tracing method for interactive architectural acoustics,” The Journal of the acoustical society of America , vol. 115, no. 2, pp. 739–756, 2004
2004
Earlier work this paper cites.
L. L. Beranek, “Analysis of Sabine and Eyring equations and their application to concert hall audience and chair absorption,” The Journal of the Acoustical Society of America , vol. 120, no. 3, pp. 1399–1410, 2006
2006
Earlier work this paper cites.
E. A. Lehmann and A. M. Johansson, “Diffuse reverberation model for efficient image-source simulation of room impulse responses,” IEEE/ACM Transactions on Audio, Speech, and Language Processing (TASLP) , vol. 18, no. 6, pp. 1429–1439, 2009
2009
Earlier work this paper cites.
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “A short-time objective intelligibility measure for time-frequency weighted noisy speech,” in Acoustics, Speech and Signal Processing (ICASSP), 2010 IEEE International Conference on . IEEE, 2010, pp. 4214–4217
2010
Earlier work this paper cites.
K. Kinoshita, M. Delcroix, T. Yoshioka, T. Nakatani, E. Habets, R. Haeb-Umbach, V. Leutnant, A. Sehr, W. Kellermann, R. Maas et al. , “The REVERB challenge: A common evaluation framework for dereverberation and recognition of reverberant speech,” in 2013 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics . IEEE, 2013, pp. 1–4
2013
Earlier work this paper cites.
J. Thiemann, N. Ito, and E. Vincent, “The diverse environments multi-channel acoustic noise database (demand): A database of multichannel environmental noise recordings,” in Proceedings of Meetings on Acoustics ICA2013 , vol. 19, no. 1. Acoustical Society of America, 2013, p. 035081
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” Advances in neural information processing systems , vol. 27, 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
2015
Cited alongside, same era.
T. N. Sainath, O. Vinyals, A. Senior, and H. Sak, “Convolutional, long short-term memory, fully connected deep neural networks,” in Acoustics, Speech and Signal Processing (ICASSP), 2015 IEEE International Conference on . IEEE, 2015, pp. 4580–4584
2015
Cited alongside, same era.
I. Szöke, M. Skácel, L. Mošner, J. Paliesek, and J. Černockỳ, “Building and evaluation of a real room impulse response dataset,” IEEE Journal of Selected Topics in Signal Processing , vol. 13, no. 4, pp. 863–876, 2019
2019
Later among the works it cites.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga et al. , “Pytorch: An imperative style, high-performance deep learning library,” Advances in neural information processing systems , vol. 32, 2019
2019
Later among the works it cites.
D. Diaz-Guerra, A. Miguel, and J. R. Beltran, “gpuRIR: A python library for room impulse response simulation with gpu acceleration,” Multimedia Tools and Applications , pp. 1–19, 2020
2020
Later among the works it cites.
Z. Tang, L. Chen, B. Wu, D. Yu, and D. Manocha, “Improving reverberant speech training using diffuse acoustic simulation,” in Acoustics, Speech and Signal Processing (ICASSP), 2020 IEEE International Conference on . IEEE, 2020, pp. 6969–6973
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. S. Williamson, Y. Wang, and D. Wang, “Complex ratio masking for monaural speech separation,” IEEE/ACM Transactions on Audio, Speech, and Language Processing (TASLP) , vol. 24, no. 3, pp. 483–492, 2015
2015
Cited alongside, same era.
Z.-h. Fu and J.-w. Li, “GPU-based image method for room impulse response calculation,” Multimedia Tools and Applications , vol. 75, no. 9, pp. 5205–5221, 2016
2016
Cited alongside, same era.
C. Kim, A. Misra, K. Chin, T. Hughes, A. Narayanan, T. N. Sainath, and M. Bacchiani, “Generation of large-scale simulated utterances in virtual rooms to train deep-neural networks for far-field speech recognition in google home,” Proc. Interspeech , pp. 379–383, 2017
2017
Cited alongside, same era.
R. Scheibler, E. Bezzam, and I. Dokmanić, “Pyroomacoustics: A python package for audio room simulation and array processing algorithms,” in Acoustics, Speech and Signal Processing (ICASSP), 1997 IEEE International Conference on . IEEE, 2018, pp. 351–355
2018
Cited alongside, same era.
2018
Cited alongside, same era.
“RIR Generator,” https://www.audiolabs-erlangen.de/fau/professor/habets/software/rir-generator
Cited in the paper.
Later among the works it cites.
2020
Later among the works it cites.
P. Masztalski, M. Matuszewski, K. Piaskowski, and M. Romaniuk, “StoRIR: Stochastic room impulse response generation for audio data augmentation,” Proc. Interspeech , pp. 2857–2861, 2020
2020
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Ratnarajah, S.-X. Zhang, M. Yu, Z. Tang, D. Manocha, and D. Yu, “FAST-RIR: Fast neural diffuse room impulse response generator,” in Acoustics, Speech and Signal Processing (ICASSP), 2022 IEEE International Conference on . IEEE, 2022, pp. 571–575
2022
Closest in time.