Fetching the paper…
Reading the bibliography…
Voice-enabled technology is quickly becoming ubiquitous, and is constituted from machine learning (ML)-enabled components such as speech recognition and voice activity detection.
The Ethnographic Interview
James P. Spradley · 1979
Earlier work this paper cites.
Naturalistic Inquiry
Yvonna S Lincoln and Egon G Guba · 1985
Earlier work this paper cites.
In-Depth Interviewing: Researching People
Victor Minichiello, Rosalie Aroni, Eric Timewell, and Loris Alexander · 1990
Earlier work this paper cites.
Grounded theory research: Procedures, canons, and evaluative criteria
Juliet M Corbin and Anselm Strauss · 1990
Earlier work this paper cites.
Bias in computer systems
Batya Friedman and Helen Nissenbaum · 1996
Earlier work this paper cites.
The politics of transcription
Mary Bucholtz · 2000
Earlier work this paper cites.
Communities of Practice and Social Learning Systems
Etienne Wenger · 2000
Earlier work this paper cites.
Variation in transcription
Mary Bucholtz · 2007
Earlier work this paper cites.
The Syntactic Atlas of the Dutch Dialects (SAND): A corpus of elicited speech and text as an online dynamic atlas
Sjef Barbiers, Leonie Cornips, and Jan Pieter Kunst · 2007
Earlier work this paper cites.
Toward information infrastructure studies: Ways of knowing in a networked environment
Geoffrey C Bowker, Karen Baker, Florence Millerand, and David Ribes · 2009
Earlier work this paper cites.
Document analysis as a qualitative research method
Glenn A Bowen · 2009
Earlier work this paper cites.
Survey Methodology , volume 561
Robert M Groves, Floyd J Fowler Jr, Mick P Couper, James M Lepkowski, Eleanor Singer, and Roger Tourangeau · 2011
Earlier work this paper cites.
Standards: Recipes for Reality
Lawrence Busch · 2011
Earlier work this paper cites.
TED-LIUM: An Automatic Speech Recognition dedicated corpus
Anthony Rousseau, Paul Deléglise, and Yannick Esteve · 2012
Earlier work this paper cites.
Coding data and interpreting text: Methods of analysis
Douglas Ezzy · 2013
Earlier work this paper cites.
Code-Switching in Conversation: Language, Interaction and Identity
Peter Auer · 2013
Earlier work this paper cites.
The Data Revolution: Big Data, Open Data, Data Infrastructures and Their Consequences
Rob Kitchin · 2014
Earlier work this paper cites.
Language: Its Structure and Use
Edward Finegan · 2014
Earlier work this paper cites.
Learning in a landscape of practice: A framework
Etienne Wenger-Trayner and Beverly Wenger-Trayner · 2014
Earlier work this paper cites.
Voice assistant application for the Serbian language
Branislav Popović, Edvin Pakoci, Nikša Jakovljević, Goran Kočiš, and Darko Pekar · 2015
Earlier work this paper cites.
Librispeech: An ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Earlier work this paper cites.
Principles of dataset versioning: Exploring the recreation/storage tradeoff
Souvik Bhattacherjee, Amit Chavan, Silu Huang, Amol Deshpande, and Aditya Parameswaran · 2015
Earlier work this paper cites.
Gartner Says Worldwide Spending on VPA-Enabled Wireless Speakers Will Top $2 Billion by 2020
Rob Van der Meulen and Amy Ann Forni · 2016
Earlier work this paper cites.
Collecting resources in sub-Saharan African languages for automatic speech recognition: A case study of wolof
Elodie Gauthier, Laurent Besacier, Sylvie Voisin, Michael Melese, and Uriel Pascal Elingui · 2016
Earlier work this paper cites.
Metadata
Marcia Lei Zeng and Jian Qin · 2016
Earlier work this paper cites.
Hardware architectures for embedded speaker recognition applications: A survey
Hasna Bouraoui, Chadlia Jerad, Anupam Chattopadhyay, and Nejib Ben Hadj-Alouane · 2017
Earlier work this paper cites.
Gender and dialect bias in YouTube’s automatic captions
Rachael Tatman · 2017
Earlier work this paper cites.
Effects of talker dialect, gender & race on accuracy of bing speech and YouTube automatic captions
Rachael Tatman and Conner Kasten · 2017
Earlier work this paper cites.
Voxceleb: A large-scale speaker identification dataset
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman · 2017
Cited alongside, same era.
LJ Speech, 2017
Keith Ito · 2017
Cited alongside, same era.
Data statements for natural language processing: Toward mitigating system bias and enabling better science
Emily M Bender and Batya Friedman · 2018
Cited alongside, same era.
What makes users trust a chatbot for customer service? An exploratory interview study
Asbjørn Følstad, Cecilie Bertinussen Nordheim, and Cato Alexander Bjørkli · 2018
Cited alongside, same era.
Research Design
John W Creswell and J David Creswell · 2018
Cited alongside, same era.
Targeting the benchmark: On methodology in current natural language processing research
David Schlangen · 2020
Later among the works it cites.
The Voice Catchers
Joseph Turow · 2021
Later among the works it cites.
Quantifying bias in automatic speech recognition
Siyuan Feng, Olya Kudina, Bence Mark Halpern, and Odette Scharenborg · 2021
Later among the works it cites.
Accented speech recognition: A survey
Arthur Hinsvark, Natalie Delworth, Miguel Del Rio, Quinten McNamara, Joshua Dong, Ryan Westerman, Michelle Huang, Joseph Palakapilly, Jennifer Drexler, Ilya Pirkin, et al · 2021
Later among the works it cites.
Datasheets for datasets
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé Iii, and Kate Crawford · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman · 2018
Cited alongside, same era.
TED-LIUM 3: Twice as much data and corpus repartition for experiments on speaker adaptation
François Hernandez, Vincent Nguyen, Sahar Ghannay, Natalia Tomashenko, and Yannick Estève · 2018
Cited alongside, same era.
Jakobovski/free-spoken-digit-dataset: V1.0.8
Zohar Jackson, César Souza, Jason Flaks, Yuxin Pan, Hereman Nicolas, and Adhish Thite · 2018
Cited alongside, same era.
The fifth ’CHiME’ Speech separation and recognition challenge: Dataset, task and baselines
Jon Barker, Shinji Watanabe, Emmanuel Vincent, and Jan Trmal · 2018
Cited alongside, same era.
Data infrastructure literacy
Jonathan Gray, Carolin Gerlitz, and Liliana Bounegru · 2018
Cited alongside, same era.
Voice assistant to remind pharmacologic treatment in elders
Manuel Jesús-Azabal, Javier Rojo, Enrique Moguel, Daniel Flores-Martin, Javier Berrocal, José García-Alonso, and Juan M Murillo · 2019
Cited alongside, same era.
Model cards for model reporting
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru · 2019
Cited alongside, same era.
Datasheets for datasets help ML engineers notice and understand ethical issues in training data
Karen L Boyd · 2021
Later among the works it cites.
Documenting computer vision datasets: An invitation to reflexive data practices
Milagros Miceli, Tianling Yang, Laurens Naudts, Martin Schuessler, Diana Serbanescu, and Alex Hanna · 2021
Later among the works it cites.
On the genealogy of machine learning datasets: A critical history of ImageNet
Emily Denton, Alex Hanna, Razvan Amironesei, Andrew Smart, and Hilary Nicole · 2021
Later among the works it cites.
Jack Bandy and Nicholas Vincent · 2021
Later among the works it cites.
Angelina McMillan-Major, Salomey Osei, Juan Diego Rodriguez, Pawan Sasanka Ammanamanchi, Sebastian Gehrmann, and Yacine Jernite · 2021
Later among the works it cites.
Ready or not, AI comes—an interview study of organizational AI readiness factors
Jan Jöhnk, Malte Weißert, and Katrin Wyrtki · 2021
Later among the works it cites.
The Coding Manual for Qualitative Researchers
Johnny Saldaña · 2021
Later among the works it cites.
Librivox - Acoustical liberation of books in the public domain., 2021
Librivox · 2021
Later among the works it cites.
AI and the everything in the whole wide world benchmark
Inioluwa Deborah Raji, Emily M Bender, Amandalynne Paullada, Emily Denton, and Alex Hanna · 2021
Later among the works it cites.
More voice, less choice: The rise of voice interfaces and the decline of open source voice, January 2021
Kathy Reid · 2021
Later among the works it cites.
An empirical study of older adult’s voice assistant use for health information seeking
Robin Brewer, Casey Pierce, Pooja Upadhyay, and Leeseul Park · 2022
Later among the works it cites.
Data2vec: A general framework for self-supervised learning in speech, vision and language
Alexei Baevski, Wei-Ning Hsu, Qiantong Xu, Arun Babu, Jiatao Gu, and Michael Auli · 2022
Later among the works it cites.
Towards measuring fairness in speech recognition: Casual conversations dataset transcriptions
Chunxi Liu, Michael Picheny, Leda Sarı, Pooja Chitkara, Alex Xiao, Xiaohui Zhang, Mark Chou, Andres Alvarado, Caner Hazirbas, and Yatharth Saraf · 2022
Later among the works it cites.
Hey ASR system! Why aren’t you more inclusive? Automatic speech recognition systems’ bias and proposed bias mitigation techniques. A literature review
Mikel K Ngueajio and Gloria Washington · 2022
Later among the works it cites.
Bias in automated speaker recognition
Wiebke Toussaint Hutiri and Aaron Yi Ding · 2022
Later among the works it cites.
Data cards: Purposeful and transparent dataset documentation for responsible AI
Mahima Pushkarna, Andrew Zaldivar, and Oddur Kjartansson · 2022
Later among the works it cites.
The model card authoring toolkit: Toward community-centered, deliberation-driven AI design
Hong Shen, Leijie Wang, Wesley H Deng, Ciell Brusse, Ronald Velgersdijk, and Haiyi Zhu · 2022
Later among the works it cites.
Interactive model cards: A human-centered approach to model documentation
Anamaria Crisan, Margaret Drouhard, Jesse Vig, and Nazneen Rajani · 2022
Later among the works it cites.
Understanding machine learning practitioners’ data documentation perceptions, needs, challenges, and desiderata
Amy K Heger, Liz B Marquis, Mihaela Vorvoreanu, Hanna Wallach, and Jennifer Wortman Vaughan · 2022
Later among the works it cites.
The values encoded in machine learning research
Abeba Birhane, Pratyusha Kalluri, Dallas Card, William Agnew, Ravit Dotan, and Michelle Bao · 2022
Later among the works it cites.
Model Cards, December 2022
Ezi Ozeani, Marisa Gerchick, and Margaret Mitchell · 2022
Later among the works it cites.
Zahra Ashktorab, Benjamin Hoover, Mayank Agarwal, Casey Dugan, Werner Geyer, Hao Bang Yang, and Mikhail Yurochkin · 2023
Closest in time.
Common Voice and Accent Choice: Data Contributors Self-Describe Their Spoken Accents in Diverse Ways
Kathy Reid and Elizabeth T. Williams · 2023
Closest in time.