Fetching the paper…
Reading the bibliography…
Audio signals are often stored and transmitted in compressed formats.
“Language Models are Few-Shot Learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“Distance between sets”
Michael Levandowsky and David Winter · 1971
Earlier work this paper cites.
“Signal compression based on models of human perception”
N. Jayant, J. Johnston and R. Safranek · 1993
Earlier work this paper cites.
“ISO/IEC 13818-3:1995 - Information technology – Generic coding of moving pictures and associated audio information – Part 3: Audio”, https://www.iso.org/standard/22991.html
1995
Earlier work this paper cites.
“ISO/IEC 13818-7:1997 - Information technology – Generic coding of moving pictures and associated audio information – Part 7: Advanced Audio Coding (AAC)”, https://www.iso.org/standard/25040.html
1997
Earlier work this paper cites.
“MP3 and AAC explained” Signa, Italy
Karlheinz Brandenburg · 1999
Earlier work this paper cites.
“The Theory Behind MP3”, 2002
Rassol Raissi · 2002
Earlier work this paper cites.
“Musical genre classification of audio signals”
G. Tzanetakis and P. Cook · 2002
Earlier work this paper cites.
“Genesis of the MP3 audio coding standard”
H.G. Musmann · 2006
Earlier work this paper cites.
“MP3 decoder in theory and practice”, 2006
Praveen Sripada · 2006
Earlier work this paper cites.
“Defeating fake-quality MP3” Princeton, NJ, USA
Rui Yang, Yun-Qing Shi and Jiwu Huang · 2009
Earlier work this paper cites.
“Detection of double MP3 compression”
Qingzhong Liu, Andrew Sung and Mengyu Qiao · 2010
Earlier work this paper cites.
“Detecting double compression of audio signal”
Rui Yang, Yun. Shi and Jiwu Huang · 2010
Cited alongside, same era.
“The Balanced Accuracy and Its Posterior Distribution” Istanbul, Turkey
Kay Brodersen, Cheng Ong, Klaas Stephan and Joachim. Buhmann · 2010
Cited alongside, same era.
“Improved detection of MP3 double compression using content-independent features” KunMing, China
Mengyu Qiao, Andrew Sung and Qingzhong Liu · 2013
Cited alongside, same era.
“Detecting double-compressed MP3 with the Same Bit-rate.”
Pengfei Ma, Rangding Wang, Diqun Yan and Chao Jin · 2014
Cited alongside, same era.
“A Huffman Table Index Based Approach to Detect Double MP3 Compression”
Pengfei Ma, Rangding Wang, Diqun Yan and Chao Jin · 2014
Cited alongside, same era.
“Detection and localization of double compression in MP3 audio tracks”
Tiziano Bianchi, Alessia De, Marco Fontani, Giovanni Rocciolo and Alessandro Piva · 2014
“Compression history detection for MP3 audio”
Diqun Yan, Rangding Wang, Jinglei Zhou, Chao Jin and Zhifeng Wang · 2018
Later among the works it cites.
“Waveglow: A flow-based generative network for speech synthesis” Brighton, UK
Ryan Prenger, Rafael Valle and Bryan Catanzaro · 2019
Later among the works it cites.
“Enabling Factorized Piano Music Modeling and Generation with the MAESTRO Dataset”
Curtis Hawthorne, Andriy Stasyuk, Adam Roberts, Ian Simon, Cheng-Zhi Huang, Sander Dieleman, Erich Elsen, Jesse Engel and Douglas Eck · 2019
Later among the works it cites.
“Fastspeech 2: Fast and high-quality end-to-end text to speech”
Yi Ren, Chenxu Hu, Xu Tan, Tao Qin, Sheng Zhao, Zhou Zhao and Tie-Yan Liu · 2020
Later among the works it cites.
“Jukebox: A generative model for music”
Prafulla Dhariwal, Heewoo Jun, Christine Payne, Jong Kim, Alec Radford and Ilya Sutskever · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
“Very deep convolutional neural network based image classification using small training sample size”
Shuying Liu and Weihong Deng · 2015
Cited alongside, same era.
“An overview of gradient descent optimization algorithms”
Sebastian Ruder · 2016
Cited alongside, same era.
“Gaussian error linear units (gelus)”
Dan Hendrycks and Kevin Gimpel · 2016
Cited alongside, same era.
“Attention is all you need”
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan Gomez, ukasz Kaiser and Illia Polosukhin · 2017
Cited alongside, same era.
“The LJ Speech Dataset”, 2017
Keith Ito and Linda Johnson · 2017
Cited alongside, same era.
“Compression Detection of Audio Waveforms Based on Stacked Autoencoders”
Da Luo, Wenqing Cheng, Huaqiang Yuan, Weiqi Luo and Zhenghui Liu · 2020
Later among the works it cites.
“An image is worth 16x16 words: Transformers for image recognition at scale”
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold and Sylvain Gelly · 2020
Later among the works it cites.
“End-to-End Object Detection with Transformers”
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov and Sergey Zagoruyko · 2020
Later among the works it cites.
“Evaluation: from precision, recall and F-measure to ROC, informedness, markedness and correlation”
David Powers · 2020
Later among the works it cites.
“Synthetic speech detection through short-term and long-term prediction traces”
C. Borrelli, P. Bestagini, F. Antonacci, A. Sarti and S. Tubaro · 2021
Later among the works it cites.
“Highly accurate protein structure prediction with AlphaFold”
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Z\’dek and Anna Potapenko · 2021
Later among the works it cites.