Fetching the paper…
Reading the bibliography…
We present a neural-network-based fast diffuse room impulse response generator (FAST-RIR) for generating room impulse responses (RIRs) for a given acoustic environment.
“Image method for efficiently simulating small-room acoustics,”
Jont B. Allen and David A. Berkley, · 1979
Earlier work this paper cites.
“Collected papers on acoustics,” 1994
Wallace Clement Sabine and M David Egan, · 1994
Earlier work this paper cites.
“Acoustical finite‐difference time‐domain simulation in a quasi‐cartesian grid,”
D. Botteldooren, · 1994
Earlier work this paper cites.
“Finite‐difference time‐domain simulation of low‐frequency room acoustic problems,”
D. Botteldooren, · 1995
Earlier work this paper cites.
“The ami meeting corpus: A pre-announcement,”
J Carletta et al., · 2005
Earlier work this paper cites.
Room acoustics
Heinrich Kuttruff, · 2009
Earlier work this paper cites.
“Generative adversarial nets,”
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio, · 2014
Earlier work this paper cites.
“Conditional generative adversarial nets,”
Mehdi Mirza and Simon Osindero, · 2014
Earlier work this paper cites.
“Three ways to adapt a CTS recognizer to unseen reverberated speech in BUT system for the ASpIRE challenge,”
Martin Karafiát, František Grézl, Lukáš Burget, Igor Szöke, and Jan Černocký, · 2015
Earlier work this paper cites.
“Overview of geometrical room acoustic modeling techniques,”
Lauri Savioja and U. Peter Svensson, · 2015
Earlier work this paper cites.
“Conditional generative adversarial networks for convolutional face generation,”
Jon Gauthier, · 2015
Earlier work this paper cites.
“The ACE challenge - corpus description and performance evaluation,”
James Eaton, Nikolay D. Gaubitch, Alastair H. Moore, and Patrick A. Naylor, · 2015
Cited alongside, same era.
“Librispeech: An ASR corpus based on public domain audio books,”
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur, · 2015
Cited alongside, same era.
“Interactive sound propagation and rendering for large multi-source scenes,”
Carl Schissler and Dinesh Manocha, · 2016
Cited alongside, same era.
“Gpu-based image method for room impulse response calculation,”
Zhong-Hua Fu and Jian-Wei Li, · 2016
Cited alongside, same era.
“Generation of Large-Scale Simulated Utterances in Virtual Rooms to Train Deep-Neural Networks for Far-Field Speech Recognition in Google Home,”
Chanwoo Kim, Ananya Misra, Kean Chin, Thad Hughes, Arun Narayanan, Tara N. Sainath, and Michiel Bacchiani, · 2017
Cited alongside, same era.
“Gansynth: Adversarial neural audio synthesis,”
Jesse H. Engel, Kumar Krishna Agrawal, Shuo Chen, Ishaan Gulrajani, Chris Donahue, and Adam Roberts, · 2019
Later among the works it cites.
“Improving reverberant speech training using diffuse acoustic simulation,”
Zhenyu Tang, Lianwu Chen, Bo Wu, Dong Yu, and Dinesh Manocha, · 2020
Later among the works it cites.
“Storir: Stochastic room impulse response generation for audio data augmentation,”
Piotr Masztalski, Mateusz Matuszewski, Karol Piaskowski, and Michal Romaniuk, · 2020
Later among the works it cites.
“Sound synthesis, propagation, and rendering: a survey,”
Shiguang Liu and Dinesh Manocha, · 2020
Later among the works it cites.
“gpurir: A python library for room impulse response simulation with gpu acceleration,”
David Diaz-Guerra, Antonio Miguel, and Jose R Beltran, · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“A study on data augmentation of reverberant speech for robust speech recognition,”
Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L. Seltzer, and Sanjeev Khudanpur, · 2017
Cited alongside, same era.
“Multichannel signal processing with deep neural networks for automatic speech recognition,”
Tara N. Sainath et al., · 2017
Cited alongside, same era.
“Training data augmentation and data selection,”
Martin Karafiát, Karel Veselý, Katerina Zmolíková, Marc Delcroix, Shinji Watanabe, Lukás Burget, Jan Honza Cernocký, and Igor Szöke, · 2017
Cited alongside, same era.
“Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks,”
Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris Metaxas, · 2017
Cited alongside, same era.
“The REVERB challenge: A benchmark task for reverberation-robust ASR techniques,”
Keisuke Kinoshita et al., · 2017
Cited alongside, same era.
“Building and evaluation of a real room impulse response dataset,”
Igor Szöke, Miroslav Skácel, Ladislav Mosner, Jakub Paliesek, and Jan Honza Cernocký, · 2019
Cited alongside, same era.
“Scene-aware far-field automatic speech recognition,” 2021
Zhenyu Tang and Dinesh Manocha, · 2021
Closest in time.
“IR-GAN: Room Impulse Response Generator for Far-Field Speech Recognition,”
Anton Ratnarajah, Zhenyu Tang, and Dinesh Manocha, · 2021
Closest in time.
“Ts-rir: Translated synthetic room impulse responses for speech augmentation,”
Anton Ratnarajah, Zhenyu Tang, and Dinesh Manocha, · 2021
Closest in time.
“Image2reverb: Cross-model reverb impulse response synthesis,”
Nikhil Singh, Jeff Mentch, Jerry Ng, Matthew Beveridge, and Iddo Drori, · 2021
Closest in time.
“Directional ASR: A new paradigm for E2E multi-speaker speech recognition with source localization,”
A.S. Subramanian et al., · 2021
Closest in time.
“Complex neural spatial filter: Enhancing multi-channel target speech separation in complex domain,”
Rongzhi Gu, Shi-Xiong Zhang, Yuexian Zou, and Dong Yu, · 2021
Closest in time.