Fetching the paper…
Reading the bibliography…
Speech-centric machine learning systems have revolutionized many leading domains ranging from transportation and healthcare to education and defense, profoundly changing how people live, work, and interact with each other.
“Csr-i (wsj0) sennheiser ldc93s6b”
John Garofolo, David Graff, Doug Paul and David Pallett · 1993
Earlier work this paper cites.
“Public-key cryptosystems based on composite degree residuosity classes”
Pascal Paillier · 1999
Earlier work this paper cites.
“Speaker verification using adapted Gaussian mixture models”
Douglas Reynolds, Thomas Quatieri and Robert Dunn · 2000
Earlier work this paper cites.
“A survey of hybrid ANN/HMM models for automatic speech recognition”
Edmondo Trentin and Marco Gori · 2001
Earlier work this paper cites.
“Calibrating noise to sensitivity in private data analysis”
Cynthia Dwork, Frank McSherry, Kobbi Nissim and Adam Smith · 2006
Earlier work this paper cites.
“A framework for secure speech recognition”
Paris Smaragdis and Madhusudana Shashanka · 2007
Earlier work this paper cites.
“IEMOCAP: Interactive emotional dyadic motion capture database”
Carlos Busso et al · 2008
Earlier work this paper cites.
“Differential privacy: A survey of results”
Cynthia Dwork · 2008
Earlier work this paper cites.
“The security of machine learning”
Marco Barreno, Blaine Nelson, Anthony Joseph and J Tygar · 2010
Earlier work this paper cites.
“The YouTube video recommendation system”
James Davidson et al · 2010
Earlier work this paper cites.
“Front-end factor analysis for speaker verification”
Najim Dehak et al · 2010
Earlier work this paper cites.
“Opensmile: the munich versatile and fast open-source audio feature extractor”
Florian Eyben, Martin Wöllmer and Björn Schuller · 2010
Earlier work this paper cites.
“Finding Difficult Speakers in Automatic Speaker Recognition”, 2011
Lara Stoll · 2011
Earlier work this paper cites.
“The Speech Accent Archive: towards a typology of English accents”
Steven Weinberger and Stephen Kunath · 2011
Earlier work this paper cites.
“Poisoning attacks against support vector machines”
Battista Biggio, Blaine Nelson and Pavel Laskov · 2012
Earlier work this paper cites.
“Fairness through awareness”
Cynthia Dwork et al · 2012
Earlier work this paper cites.
“Privacy-preserving speaker verification and identification using gaussian mixture models”
Manas Pathak and Bhiksha Raj · 2012
Earlier work this paper cites.
“TED-LIUM: an Automatic Speech Recognition dedicated corpus.”
Anthony Rousseau, Paul Deléglise and Yannick Esteve · 2012
Earlier work this paper cites.
“Evasion attacks against machine learning at test time”
Battista Biggio et al · 2013
Earlier work this paper cites.
“Intriguing properties of neural networks”
Christian Szegedy et al · 2013
Earlier work this paper cites.
“Explaining and harnessing adversarial examples”
Ian Goodfellow, Jonathon Shlens and Christian Szegedy · 2014
Earlier work this paper cites.
“Deep speech: Scaling up end-to-end speech recognition”
Awni Hannun et al · 2014
Earlier work this paper cites.
“Automatic language identification using deep neural networks”
Ignacio Lopez-Moreno et al · 2014
Earlier work this paper cites.
“Sound shredding: Privacy preserved audio sensing”
Sumeet Kumar et al · 2015
Earlier work this paper cites.
“Using machine teaching to identify optimal training-set attacks on machine learners”
Shike Mei and Xiaojin Zhu · 2015
Earlier work this paper cites.
“Librispeech: an asr corpus based on public domain audio books”
Vassil Panayotov, Guoguo Chen, Daniel Povey and Sanjeev Khudanpur · 2015
Earlier work this paper cites.
“Machine bias”
Julia Angwin, Jeff Larson, Surya Mattu and Lauren Kirchner · 2016
Earlier work this paper cites.
“Equality of opportunity in supervised learning”
Moritz Hardt, Eric Price and Nati Srebro · 2016
Earlier work this paper cites.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Earlier work this paper cites.
“Mining the spoken wikipedia for speech data and beyond”
Arne Köhn, Florian Stegen and Timo Baumann · 2016
Earlier work this paper cites.
“Mastering the game of Go with deep neural networks and tree search”
David Silver et al · 2016
Earlier work this paper cites.
“Targeted backdoor attacks on deep learning systems using data poisoning”
Xinyun Chen et al · 2017
Earlier work this paper cites.
“Parseval networks: Improving robustness to adversarial examples”
Moustapha Cisse et al · 2017
Earlier work this paper cites.
“Algorithmic Bias in Autonomous Systems.”
David Danks and Alex London · 2017
Earlier work this paper cites.
“Crafting adversarial examples for speech paralinguistics applications”
Yuan Gong and Christian Poellabauer · 2017
Earlier work this paper cites.
“Badnets: Identifying vulnerabilities in the machine learning model supply chain”
Tianyu Gu, Brendan Dolan-Gavitt and Siddharth Garg · 2017
Earlier work this paper cites.
“Counterfactual fairness”
Matt Kusner, Joshua Loftus, Chris Russell and Ricardo Silva · 2017
Earlier work this paper cites.
“Trojaning attack on neural networks”, 2017
Yingqi Liu et al · 2017
Earlier work this paper cites.
“Towards deep learning models resistant to adversarial attacks”
Aleksander Madry et al · 2017
Earlier work this paper cites.
“Communication-efficient learning of deep networks from decentralized data”
Brendan McMahan et al · 2017
Earlier work this paper cites.
“Towards poisoning of deep learning algorithms with back-gradient optimization”
Luis Muñoz-Gonzãlez et al · 2017
Earlier work this paper cites.
“Voxceleb: a large-scale speaker identification dataset”
Arsha Nagrani, Joon Chung and Andrew Zisserman · 2017
Earlier work this paper cites.
“No bot expects the DeepCAPTCHA! Introducing immutable adversarial examples, with applications to CAPTCHA generation”
Margarita Osadchy et al · 2017
Earlier work this paper cites.
“The 2016 NIST Speaker Recognition Evaluation”
Seyed Sadjadi et al · 2017
Earlier work this paper cites.
“Membership inference attacks against machine learning models”
Reza Shokri, Marco Stronati, Congzheng Song and Vitaly Shmatikov · 2017
Earlier work this paper cites.
“Certified defenses for data poisoning attacks”
Jacob Steinhardt, Pang Koh and Percy Liang · 2017
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“The eu general data protection regulation (gdpr)”
Paul Voigt and Axel Von · 2017
Earlier work this paper cites.
“Generative poisoning attack method against neural networks”
Chaofei Yang, Qing Wu, Hai Li and Yiran Chen · 2017
Earlier work this paper cites.
“Efficient defenses against adversarial attacks”
Valentina Zantedeschi, Maria-Irina Nicolae and Ambrish Rawat · 2017
Earlier work this paper cites.
“mixup: Beyond empirical risk minimization”
Hongyi Zhang, Moustapha Cisse, Yann Dauphin and David Lopez-Paz · 2017
Earlier work this paper cites.
“Did you hear that? adversarial examples against automatic speech recognition”
Moustafa Alzantot, Bharathan Balaji and Mani Srivastava · 2018
Earlier work this paper cites.
“Exploring siamese neural network architectures for preserving speaker identity in speech emotion classification”
Priya Arora and Theodora Chaspari · 2018
Earlier work this paper cites.
“The fifth ’CHiME’ speech separation and recognition challenge: dataset, task and baselines”
Jon Barker, Shinji Watanabe, Emmanuel Vincent and Jan Trmal · 2018
Earlier work this paper cites.
“Expanding the reach of federated learning by reducing client resource requirements”
Sebastian Caldas, Jakub Konečny, H McMahan and Ameet Talwalkar · 2018
Earlier work this paper cites.
“Audio adversarial examples: Targeted attacks on speech-to-text”
Nicholas Carlini and David Wagner · 2018
Earlier work this paper cites.
“Detecting backdoor attacks on deep neural networks by activation clustering”
Bryant Chen et al · 2018
Earlier work this paper cites.
“Voxceleb2: Deep speaker recognition”
Joon Chung, Arsha Nagrani and Andrew Zisserman · 2018
Earlier work this paper cites.
“Bert: Pre-training of deep bidirectional transformers for language understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Earlier work this paper cites.
“Tiles audio recorder: an unobtrusive wearable solution to track audio activity”
Tiantian Feng et al · 2018
Earlier work this paper cites.
Shreya Khare, Rahul Aralikatte and Senthil Mani · 2018
Earlier work this paper cites.
“Defense against adversarial attacks using high-level representation guided denoiser”
Fangzhou Liao et al · 2018
Earlier work this paper cites.
“Fine-pruning: Defending against backdooring attacks on deep neural networks”
Kang Liu, Brendan Dolan-Gavitt and Siddharth Garg · 2018
Earlier work this paper cites.
“Deep learning for healthcare: review, opportunities and challenges”
Riccardo Miotto et al · 2018
Earlier work this paper cites.
“Meld: A multimodal multi-party dataset for emotion recognition in conversations”
Soujanya Poria et al · 2018
Earlier work this paper cites.
“Hidebehind: Enjoy voice input with voiceprint unclonability and anonymity”
Jianwei Qian et al · 2018
Earlier work this paper cites.
“Towards privacy-preserving speech data publishing”
Jianwei Qian et al · 2018
Earlier work this paper cites.
“Privacy-preserving iVector-based speaker verification”
Yogachandran Rahulamathavan et al · 2018
Earlier work this paper cites.
“Automatic speaker, age-group and gender identification from children’s speech”
Saeid Safavi, Martin Russell and Peter Jančovič · 2018
Earlier work this paper cites.
“Adversarial attacks against automatic speech recognition systems via psychoacoustic hiding”
Lea Schönherr et al · 2018
Cited alongside, same era.
“Poison frogs! targeted clean-label poisoning attacks on neural networks”
Ali Shafahi et al · 2018
Cited alongside, same era.
“X-vectors: Robust dnn embeddings for speaker recognition”
David Snyder et al · 2018
Cited alongside, same era.
“Domain adversarial training for accented speech recognition”
Sining Sun et al · 2018
Cited alongside, same era.
“Spectral signatures in backdoor attacks”
Brandon Tran, Jerry Li and Aleksander Madry · 2018
Cited alongside, same era.
“Adversarial learning of raw speech features for domain invariant speech recognition”
Aditay Tripathi, Aanchan Mohan, Saket Anand and Maneesh Singh · 2018
“Cyclic defense gan against speech adversarial attacks”
Mohammad Esmaeilpour, Patrick Cardinal and Alessandro Koerich · 2021
Later among the works it cites.
“Attribute inference attack of speech emotion recognition in federated learning settings”
Tiantian Feng et al · 2021
Later among the works it cites.
“Privacy and Utility Preserving Data Transformation for Speech Emotion Recognition”
Tiantian Feng and Shrikanth Narayanan · 2021
Later among the works it cites.
“Training speech recognition models with federated learning: A quality/cost framework”
Dhruv Guliani, Françoise Beaufays and Giovanni Motta · 2021
Later among the works it cites.
“A Computational Tool to Study Vocal Participation of Women in UN-ITU Meetings”
Rajat Hebbar et al · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Fairness definitions explained”
Sahil Verma and Julia Rubin · 2018
Cited alongside, same era.
“Speech commands: A dataset for limited-vocabulary speech recognition”
Pete Warden · 2018
Cited alongside, same era.
“Robust audio adversarial example for a physical attack”
Hiromu Yakura and Jun Sakuma · 2018
Cited alongside, same era.
“Characterizing audio adversarial examples using temporal dependency”
Zhuolin Yang, Bo Li, Pin-Yu Chen and Dawn Song · 2018
Cited alongside, same era.
“Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph”
AmirAli Zadeh et al · 2018
Cited alongside, same era.
“Emotionless: Privacy-preserving speech analysis for voice assistants”
Ranya Aloufi, Hamed Haddadi and David Boyle · 2019
Cited alongside, same era.
Shehzeen Hussain et al · 2021
Later among the works it cites.
“A review of multimodal image matching: Methods and applications”
Xingyu Jiang et al · 2021
Later among the works it cites.
“When machine learning meets privacy: A survey and outlook”
Bo Liu et al · 2021
Later among the works it cites.
“Automatic speech recognition: a survey”
Mishaim Malik, Muhammad Malik, Khawar Mehmood and Imran Makhdoom · 2021
Later among the works it cites.
“A survey on bias and fairness in machine learning”
Ninareh Mehrabi et al · 2021
Later among the works it cites.
“"I don’t Think These Devices are Very Culturally Sensitive." - Impact of Automated Speech Recognition Errors on African Americans”
Zion Mengesha et al · 2021
Later among the works it cites.
“Cross-silo federated training in the cloud with diversity scaling and semi-supervised learning”
Kishore Nandury, Anand Mohan and Frederick Weber · 2021
Later among the works it cites.
“Adversarial Disentanglement of Speaker Representation for Attribute-Driven Privacy Preservation”
Paul-Gauthier Noé et al · 2021
Later among the works it cites.
“In defense of pseudo-labeling: An uncertainty-aware pseudo-label selection framework for semi-supervised learning”
Mamshad Rizve, Kevin Duarte, Yogesh Rawat and Mubarak Shah · 2021
Later among the works it cites.
“Evaluating the Vulnerability of End-to-End Automatic Speech Recognition Models to Membership Inference Attacks”
Muhammad. Shah et al · 2021
Later among the works it cites.
“A survey on automatic multimodal emotion recognition in the wild”
Garima Sharma and Abhinav Dhall · 2021
Later among the works it cites.
“Spoofing Speaker Verification With Voice Style Transfer And Reconstruction Loss”
Thomas Thebaud, Gaël Le and Anthony Larcher · 2021
Later among the works it cites.
“SVEva Fair: A Framework for Evaluating Fairness in Speaker Verification”
Wiebke Toussaint and Aaron Ding · 2021
Later among the works it cites.
“Voxpopuli: A large-scale multilingual speech corpus for representation learning, semi-supervised learning and interpretation”
Changhan Wang et al · 2021
Later among the works it cites.
“Multiview pseudo-labeling for semi-supervised learning from video”
Bo Xiong, Haoqi Fan, Kristen Grauman and Christoph Feichtenhofer · 2021
Later among the works it cites.
“To be robust or to be fair: Towards fairness in adversarial training”
Han Xu et al · 2021
Later among the works it cites.
“Federated learning in ASR: Not as easy as you think”
Wentao Yu et al · 2021
Later among the works it cites.
“Seen and unseen emotional style transfer for voice conversion with a new emotional speech dataset”
Kun Zhou, Berrak Sisman, Rui Liu and Haizhou Li · 2021
Later among the works it cites.
“Evaluating the covid-19 identification resnet (cider) on the interspeech covid-19 from audio challenges”
Alican Akman et al · 2022
Closest in time.
“Extracting Targeted Training Data from ASR Models, and How to Mitigate It”
Ehsan Amid et al · 2022
Closest in time.
“A Study of Gender Impact in Self-supervised Models for Speech-to-Text Systems”
Marcely Boito, Laurent Besacier, Natalia. Tomashenko and Yannick Estève · 2022
Closest in time.
“VoxSRC 2021: The Third VoxCeleb Speaker Recognition Challenge”
Andrew Brown et al · 2022
Closest in time.
“Robust Federated Learning Against Adversarial Attacks for Speech Emotion Recognition”
Yi Chang et al · 2022
Closest in time.
“Exploring racial and gender disparities in voice biometrics”
Xingyu Chen, Zhengxiong Li, Srirangaraj Setlur and Wenyao Xu · 2022
Closest in time.
“Personal voice assistant security and privacy—a survey”
Peng Cheng and Utz Roedig · 2022
Closest in time.
“ILASR: privacy-preserving incremental learning for automatic speech recognition at production scale”
Gopinath Chennupati et al · 2022
Closest in time.
“Privacy-preserving Speech-based Depression Diagnosis via Federated Learning”
Yue Cui et al · 2022
Closest in time.
“Toward Fairness in Speech Recognition: Discovery and mitigation of performance disparities”
Pranav Dheram et al · 2022
Closest in time.
“HEKWS: Privacy-Preserving Convolutional Neural Network-based Keyword Spotting with a Ciphertext Packing Technique”
Daniel Elworth and Sunwoong Kim · 2022
Closest in time.
“Enhancing Privacy Through Domain Adaptive Noise Injection For Speech Emotion Recognition”
Tiantian Feng, Hanieh Hashemi, Murali Annavaram and Shrikanth Narayanan · 2022
Closest in time.
Tiantian Feng and Shrikanth Narayanan · 2022
Closest in time.
“User-Level Differential Privacy against Attribute Inference Attack of Speech Emotion Recognition on Federated Learning”
Tiantian Feng, Raghuveer Peri and Shrikanth Narayanan · 2022
Closest in time.
“WaveFuzz: A Clean-Label Poisoning Attack to Protect Your Voice”
Yunjie Ge et al · 2022
Closest in time.
“Overcoming the pitfalls and perils of algorithms: A classification of machine learning biases and mitigation methods”
Benjamin van Giffen, Dennis Herhausen and Tobias Fahse · 2022
Closest in time.
“Enabling on-device training of speech recognition models with federated dropout”
Dhruv Guliani et al · 2022
Closest in time.
“Production federated keyword spotting via distillation, filtering, and joint federated-centralized training”
Andrew Hard et al · 2022
Closest in time.
“Detecting Unintended Memorization in Language-Model-Fused ASR”
W. Huang, Steve Chien, Om Thakkar and Rajiv Mathews · 2022
Closest in time.
“Adversarial Attacks on Speech Recognition Systems for Mission-Critical Applications: A Survey”
Ngoc Huynh et al · 2022
Closest in time.
“Adversarial Reweighting for Speaker Verification Fairness”
Minho Jin et al · 2022
Closest in time.
“Towards Measuring Fairness in Speech Recognition: Casual Conversations Dataset Transcriptions”
Chunxi Liu et al · 2022
Closest in time.
“Trustworthy ai: A computational perspective”
Haochen Liu et al · 2022
Closest in time.
“Preventing sensitive-word recognition using self-supervised learning to preserve user-privacy for automatic speech recognition”
Yuchen Liu, Apu Kapadia and Donald Williamson · 2022
Closest in time.
“Model-Based Approach for Measuring the Fairness in ASR”
Zhe Liu, Irina-Elena Veliche and Fuchun Peng · 2022
Closest in time.
“Attacker Attribution of Audio Deepfakes”
Nicolas Müller, Franziska Diekmann and Jennifer Williams · 2022
Closest in time.
Raghuveer Peri, Krishna Somandepalli and Shrikanth Narayanan · 2022
Closest in time.
“Are disentangled representations all you need to build speaker anonymization systems?”
Champion Pierre, Anthony Larcher and Denis Jouvet · 2022
Closest in time.
“Robust speech recognition via large-scale weak supervision”
Alec Radford et al · 2022
Closest in time.
“AequeVox: Automated Fairness Testing of Speech Recognition Systems”
Sai Rajan, Sakshi Udeshi and Sudipta Chattopadhyay · 2022
Closest in time.
“Sotto Voce: Federated Speech Recognition with Differential Privacy Guarantees”
Michael Shoemate et al · 2022
Closest in time.
“Generating gender-ambiguous voices for privacy-preserving speech recognition”
Dimitrios Stoidis and Andrea Cavallaro · 2022
Closest in time.
“Privacy attacks for automatic speech recognition acoustic models in a federated learning framework”
Natalia Tomashenko et al · 2022
Closest in time.
“The VoicePrivacy 2022 Challenge Evaluation Plan”
Natalia Tomashenko et al · 2022
Closest in time.
“The VoicePrivacy 2020 Challenge: Results and findings”
Natalia Tomashenko et al · 2022
Closest in time.
“Membership Inference Attacks Against Self-supervised Speech Models”
Wei-Cheng Tseng, Wei-Tsung Kao and Hung-yi Lee · 2022
Closest in time.
“Privacy-preserving Speech Emotion Recognition through Semi-Supervised Federated Learning”
Vasileios Tsouvalas, Tanir Ozcelebi and Nirvana Meratnia · 2022
Closest in time.
“Dawn of the transformer era in speech emotion recognition: closing the valence gap”
Johannes Wagner et al · 2022
Closest in time.
“Partial variable training for efficient on-device federated learning”
Tien-Ju Yang, Dhruv Guliani, Françoise Beaufays and Giovanni Motta · 2022
Closest in time.
“TILES-2019: A longitudinal physiologic and behavioral data set of medical residents in an intensive care unit”
Joanna Yau et al · 2022
Closest in time.
“FedAudio: A Federated Learning Benchmark for Audio Tasks”
Tuo Zhang et al · 2022
Closest in time.
“Keyword Spotting in the Homomorphic Encrypted Domain Using Deep Complex-Valued CNN”
Peijia Zheng, Zhiwei Cai, Huicong Zeng and Jiwu Huang · 2022
Closest in time.
“Decoupled Federated Learning for ASR with Non-IID Data”
Han Zhu et al · 2022
Closest in time.
“Mel frequency spectral domain defenses against adversarial attacks on speech recognition systems”
Nicholas Mehlman, Anirudh Sreeram, Raghuveer Peri and Shrikanth Narayanan · 2023
Closest in time.
“VoiceGuard: Secure and Private Speech Processing”
Ferdinand Brasser et al · 2032
Closest in time.
“Exploring hashing and cryptonet based approaches for privacy-preserving speech emotion recognition”
Miguel Dias, Alberto Abad and Isabel Trancoso · 2061
Closest in time.