Fetching the paper…
Reading the bibliography…
Accurate estimation of Room Impulse Response (RIR), which captures an environment's acoustic properties, is important for speech processing and AR/VR applications.
Integrated‐impulse method measuring sound decay without using impulses
M. R. Schroeder · 1979
Earlier work this paper cites.
Computer‐generated pulse signal applied for sound measurement
Nobuharu Aoshima · 1981
Earlier work this paper cites.
Inverse filtering of room acoustics
M. Miyoshi and Y. Kaneda · 1988
Earlier work this paper cites.
Boundary element methods in acoustics
Carlos Alberto Brebbia and Robert D. Ciskowski · 1991
Earlier work this paper cites.
Auditorium Acoustics and Architectural Design
Michael Barron and Timothy J. Foulkes · 1994
Earlier work this paper cites.
Simultaneous echo cancellation and car noise suppression employing a microphone array
M. Dahl, I. Claesson, and S. Nordebo · 1997
Earlier work this paper cites.
Analysis of noise reduction and dereverberation techniques based on microphone arrays with postfiltering
C. Marro, Y. Mahieux, and K.U. Simmer · 1998
Earlier work this paper cites.
A multi-microphone signal subspace approach for speech enhancement
F. Jabloun and B. Champagne · 2001
Earlier work this paper cites.
Estimation of modal decay parameters from noisy response measurements
Matti Karjalainen, Poju Ansalo, Aki Mäkivirta, Timo Peltonen, and Vesa Välimäki · 2002
Earlier work this paper cites.
The ami meeting corpus: A pre-announcement
Jean Carletta, Simone Ashby, Sebastien Bourban, Mike Flynn, Mael Guillemot, Thomas Hain, Jaroslav Kadlec, Vasilis Karaiskos, Wessel Kraaij, Melissa Kronenthal, Guillaume Lathoud, Mike Lincoln, Agnes Lisowska, Iain McCowan, Wilfried Post, Dennis Reidsma, and Pierre Wellner · 2005
Earlier work this paper cites.
New Method of Measuring Reverberation Time
M. R. Schroeder · 2005
Earlier work this paper cites.
Room impulse response estimation using sparse online prediction and absolute loss
K. Crammer and D.D. Lee · 2006
Earlier work this paper cites.
Bayesian regularization and nonnegative deconvolution for room impulse response estimation
Yuanqing Lin and D.D. Lee · 2006
Earlier work this paper cites.
A review of vector quantization techniques
A. Vasuki and P.T. Vanathi · 2006
Earlier work this paper cites.
Advancements in impulse response measurements by sine sweeps
Angelo Farina · 2007
Earlier work this paper cites.
Virtual reality system with integrated sound field simulation and reproduction
Tobias Lentz, Dirk Schröder, Michael Vorländer, and Ingo Assenmacher · 2007
Earlier work this paper cites.
Speech dereverberation based on variance-normalized delayed linear prediction
Tomohiro Nakatani, Takuya Yoshioka, Keisuke Kinoshita, Masato Miyoshi, and Biing-Hwang Juang · 2010
Earlier work this paper cites.
Speech Dereverberation
Patrick A. Naylor and Nikolay D. Gaubitch · 2010
Earlier work this paper cites.
Precomputed wave simulation for real-time sound propagation of dynamic sources in complex scenes
Nikunj Raghuvanshi, John Snyder, Ravish Mehra, Ming Lin, and Naga Govindaraju · 2010
Earlier work this paper cites.
Gsound: Interactive sound propagation for games
Carl Schissler and Dinesh Manocha · 2011
Earlier work this paper cites.
Guided multiview ray tracing for fast auralization
Micah Taylor, Anish Chandak, Qi Mo, Christian Lauterbach, Carl Schissler, and Dinesh Manocha · 2012
Earlier work this paper cites.
Wave-based sound propagation in large open scenes using an equivalent source formulation
Ravish Mehra, Nikunj Raghuvanshi, Lakulish Antani, Anish Chandak, Sean Curtis, and Dinesh Manocha · 2013
Earlier work this paper cites.
Wave-ray coupling for interactive sound propagation in large complex scenes
Hengchin Yeh, Ravish Mehra, Zhimin Ren, Lakulish Antani, Dinesh Manocha, and Ming Lin · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Librispeech: An asr corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Earlier work this paper cites.
Interactive sound propagation with bidirectional path tracing
Chunxiao Cao, Zhong Ren, Carl Schissler, Dinesh Manocha, and Kun Zhou · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
A summary of the reverb challenge: State-of-the-art and remaining challenges in reverberant speech processing research
Keisuke Kinoshita, Marc Delcroix, Sharon Gannot, Emanuël Habets, Reinhold Haeb-Umbach, Walter Kellermann, Volker Leutnant, Roland Maas, Tomohiro Nakatani, Bhiksha Raj, Armin Sehr, and Takuya Yoshioka · 2016
Earlier work this paper cites.
Small-rooms dedicated to music: From room response analysis to acoustic design
Lorenzo Rizzi, Gabriele Ghelfi, and Maurizio Santini · 2016
Earlier work this paper cites.
Interactive sound propagation and rendering for large multi-source scenes
Carl Schissler and Dinesh Manocha · 2016
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niessner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang · 2017
Cited alongside, same era.
Generative Adversarial Network-Based Postfilter for STFT Spectrograms
Takuhiro Kaneko, Shinji Takaki, Hirokazu Kameoka, and Junichi Yamagishi · 2017
Cited alongside, same era.
Neural discrete representation learning
Aaron van den Oord, Oriol Vinyals, and koray kavukcuoglu · 2017
Cited alongside, same era.
Steam audio, 2018
2018
Cited alongside, same era.
Looking to listen at the cocktail party: A speaker-independent audio-visual model for speech separation
Ariel Ephrat, Inbar Mosseri, Oran Lang, Tali Dekel, Kevin Wilson, Avinatan Hassidim, William T. Freeman, and Michael Rubinstein · 2018
Cited alongside, same era.
Speech dereverberation using fully convolutional networks
Ori Ernst, Shlomo E. Chazan, Sharon Gannot, and Jacob Goldberger · 2018
Filtered noise shaping for time domain room impulse response estimation from reverberant speech
Christian J. Steinmetz, Vamsi Krishna Ithapu, and Paul Calamia · 2021
Later among the works it cites.
Soundspaces 2.0: A simulation platform for visual-acoustic learning
Changan Chen, Carl Schissler, Sanchit Garg, Philip Kobernik, Alexander Clegg, Paul Calamia, Dhruv Batra, Philip W Robinson, and Kristen Grauman · 2022
Later among the works it cites.
High fidelity neural audio compression
Alexandre Défossez, Jade Copet, Gabriel Synnaeve, and Yossi Adi · 2022
Later among the works it cites.
Sparse modeling of the early part of noisy room impulse responses with sparse bayesian learning
Maozhong Fu, Jesper Rindom Jensen, Yuhan Li, and Mads Græsbøll Christensen · 2022
Later among the works it cites.
Cortical adaptation to sound reverberation
Aleksandar Z Ivanov, Andrew J King, Ben DB Willmore, Kerry MM Walker, and Nicol S Harper · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Visual Speech Enhancement
Aviv Gabbay, Asaph Shamir, and Shmuel Peleg · 2018
Cited alongside, same era.
Audio-visual speech enhancement using multimodal deep convolutional neural networks
Jen-Cheng Hou, Syu-Siang Wang, Ying-Hui Lai, Yu Tsao, Hsiu-Wen Chang, and Hsin-Min Wang · 2018
Cited alongside, same era.
Absorption coefficient database, 2018
Christoph Kling · 2018
Cited alongside, same era.
Acoustic classification and optimization for multi-modal rendering of real-world scenes
Carl Schissler, Christian Loftin, and Dinesh Manocha · 2018
Cited alongside, same era.
Oculus spatializer, 2019
2019
Cited alongside, same era.
Microsoft project acoustics, 2019
2019
Cited alongside, same era.
Complex-Valued Time-Frequency Self-Attention for Speech Dereverberation
Vinay Kothapally and John H.L. Hansen · 2022
Later among the works it cites.
VoiceFixer: A Unified Framework for High-Fidelity Speech Restoration
Haohe Liu, Xubo Liu, Qiuqiang Kong, Qiao Tian, Yan Zhao, DeLiang Wang, Chuanzeng Huang, and Yuxuan Wang · 2022
Later among the works it cites.
Sound Synthesis, Propagation, and Rendering
Shiguang Liu and Dinesh Manocha · 2022
Later among the works it cites.
Learning neural acoustic fields
Andrew Luo, Yilun Du, Michael Tarr, Josh Tenenbaum, Antonio Torralba, and Chuang Gan · 2022
Later among the works it cites.
Few-shot audio-visual learning of environment acoustics
Sagnik Majumder, Changan Chen, Ziad Al-Halah, and Kristen Grauman · 2022
Later among the works it cites.
Interactive and Immersive Auralization
Nikunj Raghuvanshi and Hannes Gamper · 2022
Later among the works it cites.
Fast-rir: Fast neural diffuse room impulse response generator
Anton Ratnarajah, Shi-Xiong Zhang, Meng Yu, Zhenyu Tang, Dinesh Manocha, and Dong Yu · 2022
Later among the works it cites.
Gwa: A large high-quality acoustic dataset for audio processing
Zhenyu Tang, Rohith Aralikatti, Anton Jeran Ratnarajah, and Dinesh Manocha · 2022
Later among the works it cites.
Audio-visual speech codecs: Rethinking audio-visual speech enhancement by re-synthesis
Karren Yang, Dejan Marković, Steven Krenn, Vasu Agrawal, and Alexander Richard · 2022
Later among the works it cites.
Soundstream: An end-to-end neural audio codec
Neil Zeghidour, Alejandro Luebs, Ahmed Omran, Jan Skoglund, and Marco Tagliasacchi · 2022
Later among the works it cites.
Audiocite.net: Livres audio gratuits mp3, 2023
2023
Closest in time.
Novel-view acoustic synthesis
Changan Chen, Alexander Richard, Roman Shapovalov, Vamsi Krishna Ithapu, Natalia Neverova, Kristen Grauman, and Andrea Vedaldi · 2023
Closest in time.
Learning audio-visual dereverberation
Changan Chen, Wei Sun, David Harwath, and Kristen Grauman · 2023
Closest in time.
Adverb: Visually guided audio dereverberation
Sanjoy Chowdhury, Sreyan Ghosh, Subhrajyoti Dasgupta, Anton Ratnarajah, Utkarsh Tyagi, and Dinesh Manocha · 2023
Closest in time.
Tag2text: Guiding vision-language model via image tagging
Xinyu Huang, Youcai Zhang, Jinyu Ma, Weiwei Tian, Rui Feng, Yuejie Zhang, Yaqian Li, Yandong Guo, and Lei Zhang · 2023
Closest in time.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick · 2023
Closest in time.
Yet another generative model for room impulse response estimation
Sungho Lee, Hyeong-Seok Choi, and Kyogu Lee · 2023
Closest in time.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Shilong Liu, Zhaoyang Zeng, Tianhe Ren, Feng Li, Hao Zhang, Jie Yang, Chunyuan Li, Jianwei Yang, Hang Su, Jun Zhu, et al · 2023
Closest in time.
Listen2Scene: Interactive material-aware binaural sound propagation for reconstructed 3D scenes
Anton Ratnarajah and Dinesh Manocha · 2023
Closest in time.
Towards improved room impulse response estimation for speech recognition
Anton Ratnarajah, Ishwarya Ananthabhotla, Vamsi Krishna Ithapu, Pablo Hoffmann, Dinesh Manocha, and Paul Calamia · 2023
Closest in time.
Self-supervised visual acoustic matching
Arjun Somayazulu, Changan Chen, and Kristen Grauman · 2023
Closest in time.
Cross-domain diffusion based speech enhancement for very noisy speech
Heming Wang and DeLiang Wang · 2023
Closest in time.
Audiodec: An open-source streaming high-fidelity neural audio codec
Yi-Chiao Wu, Israel D. Gebru, Dejan Marković, and Alexander Richard · 2023
Closest in time.
Hifi-codec: Group-residual vector quantization for high fidelity audio codec
Dongchao Yang, Songxiang Liu, Rongjie Huang, Jinchuan Tian, Chao Weng, and Yuexian Zou · 2023
Closest in time.