Fetching the paper…
Reading the bibliography…
This paper presents the machine learning architecture of the Snips Voice Platform, a software solution to perform Spoken Language Understanding on microprocessors typical of IoT devices.
Estimation of probabilities from sparse data for the language model component of a speech recognizer
Slava Katz · 1987
Earlier work this paper cites.
Class-based n-gram models of natural language
Peter F Brown, Peter V Desouza, Robert L Mercer, Vincent J Della Pietra, and Jenifer C Lai · 1992
Earlier work this paper cites.
Finding consensus in speech recognition: word error minimization and other applications of confusion networks
Lidia Mangu, Eric Brill, and Andreas Stolcke · 2000
Earlier work this paper cites.
Conditional random fields: Probabilistic models for segmenting and labeling sequence data
John Lafferty, Andrew McCallum, and Fernando CN Pereira · 2001
Earlier work this paper cites.
Privacy by design—principles of privacy-aware ubiquitous systems
Marc Langheinrich · 2001
Earlier work this paper cites.
Weighted finite-state transducers in speech recognition
Mehryar Mohri, Fernando Pereira, and Michael Riley · 2001
Earlier work this paper cites.
Confidence measures for speech recognition: A survey
Hui Jiang · 2005
Earlier work this paper cites.
Semi-supervised learning for natural language
Percy Liang · 2005
Earlier work this paper cites.
Beyond asr 1-best: Using word confusion networks in spoken language understanding
Dilek Hakkani-Tür, Frédéric Béchet, Giuseppe Riccardi, and Gokhan Tur · 2006
Earlier work this paper cites.
Discriminative models for spoken language understanding
Ye-Yi Wang and Alex Acero · 2006
Earlier work this paper cites.
Openfst: A general and efficient weighted finite-state transducer library
Cyril Allauzen, Michael Riley, Johan Schalkwyk, Wojciech Skut, and Mehryar Mohri · 2007
Earlier work this paper cites.
Generative and discriminative algorithms for spoken language understanding
Christian Raymond and Giuseppe Riccardi · 2007
Earlier work this paper cites.
Age and gender recognition for telephone applications based on gmm supervectors and support vector machines
Tobias Bocklet, Andreas Maier, Josef G Bauer, Felix Burkhardt, and Elmar Noth · 2008
Earlier work this paper cites.
Speech recognition with weighted finite-state transducers
Mehryar Mohri, Fernando Pereira, and Michael Riley · 2008
Earlier work this paper cites.
Cheap and fast—but is it good?: evaluating non-expert annotations for natural language tasks
Rion Snow, Brendan O’Connor, Daniel Jurafsky, and Andrew Y Ng · 2008
Earlier work this paper cites.
Behavioural biometrics: a survey and classification
Roman V Yampolskiy and Venu Govindaraju · 2008
Earlier work this paper cites.
A generalized composition algorithm for weighted finite-state transducers
Cyril Allauzen, Michael Riley, and Johan Schalkwyk · 2009
Earlier work this paper cites.
Privacy by design
Ann Cavoukian · 2009
Earlier work this paper cites.
Development of an automated speech recognition interface for personal emergency response systems
Melinda Hamill, Vicky Young, Jennifer Boger, and Alex Mihailidis · 2009
Earlier work this paper cites.
Filters for efficient composition of weighted finite-state transducers
Cyril Allauzen, Michael Riley, and Johan Schalkwyk · 2010
Earlier work this paper cites.
Query language modeling for voice search
Ciprian Chelba, Johan Schalkwyk, Thorsten Brants, Vida Ha, Boulos Harb, Will Neveitt, Carolina Parada, and Peng Xu · 2010
Cited alongside, same era.
“your word is my command”: Google search by voice: a case study
Johan Schalkwyk, Doug Beeferman, Françoise Beaufays, Bill Byrne, Ciprian Chelba, Mike Cohen, Maryam Kamvar, and Brian Strope · 2010
Cited alongside, same era.
Distant speech recognition in a smart home: Comparison of several multisource asrs in realistic conditions
Benjamin Lecouteux, Michel Vacher, and François Portet · 2011
Cited alongside, same era.
The kaldi speech recognition toolkit
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al · 2011
Cited alongside, same era.
Minimum bayes risk decoding and system combination based on a recursion for edit distance
Haihua Xu, Daniel Povey, Lidia Mangu, and Jie Zhu · 2011
Cited alongside, same era.
Using recurrent neural networks for slot filling in spoken language understanding
Grégoire Mesnil, Yann Dauphin, Kaisheng Yao, Yoshua Bengio, Li Deng, Dilek Hakkani-Tur, Xiaodong He, Larry Heck, Gokhan Tur, Dong Yu, et al · 2015
Later among the works it cites.
Librispeech: an asr corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Later among the works it cites.
A time delay neural network architecture for efficient modeling of long temporal contexts
Vijayaditya Peddinti, Daniel Povey, and Sanjeev Khudanpur · 2015
Later among the works it cites.
End-to-end attention-based large vocabulary speech recognition
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Philemon Brakel, and Yoshua Bengio · 2016
Later among the works it cites.
How to add word classes to the kaldi speech recognition toolkit
Axel Horndasch, Caroline Kaufhold, and Elmar Nöth · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Calibration of confidence measures in speech recognition
Dong Yu, Jinyu Li, and Li Deng · 2011
Cited alongside, same era.
Large scale language modeling in automatic speech recognition
Ciprian Chelba, Dan Bikel, Maria Shugrina, Patrick Nguyen, and Shankar Kumar · 2012
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al · 2012
Cited alongside, same era.
Acoustic modeling using deep belief networks
Abdel-rahman Mohamed, George E Dahl, and Geoffrey Hinton · 2012
Cited alongside, same era.
Multimodal interaction in the car: combining speech and gestures on the steering wheel
Bastian Pfleging, Stefan Schneegass, and Albrecht Schmidt · 2012
Cited alongside, same era.
Crowdsourcing research opportunities: lessons from natural language processing
Marta Sabou, Kalina Bontcheva, and Arno Scharl · 2012
Cited alongside, same era.
Named entity recognition: Exploring features
Maksim Tkachenko and Andrey Simanovsky · 2012
Cited alongside, same era.
Phonetisaurus: Exploring grapheme-to-phoneme conversion with joint n-gram models in the wfst framework
Josef Robert Novak, Nobuaki Minematsu, and Keikichi Hirose · 2016
Later among the works it cites.
Purely sequence-trained neural networks for asr based on lattice-free mmi
Daniel Povey, Vijayaditya Peddinti, Daniel Galvez, Pegah Ghahremani, Vimal Manohar, Xingyu Na, Yiming Wang, and Sanjeev Khudanpur · 2016
Later among the works it cites.
On the compression of recurrent neural networks with an application to lvcsr acoustic modeling for embedded speech recognition
Rohit Prabhavalkar, Ouais Alsharif, Antoine Bruguier, and Lan McGraw · 2016
Later among the works it cites.
Achieving human parity in conversational speech recognition
Wayne Xiong, Jasha Droppo, Xuedong Huang, Frank Seide, Mike Seltzer, Andreas Stolcke, Dong Yu, and Geoffrey Zweig · 2016
Later among the works it cites.
Rasa: Open source language understanding and dialogue management
Tom Bocklisch, Joey Faulker, Nick Pawlowski, and Alan Nichol · 2017
Later among the works it cites.
Evaluating natural language understanding services for conversational question answering systems
Daniel Braun, Adrian Hernandez-Mendez, Florian Matthes, and Manfred Langen · 2017
Later among the works it cites.
Language, engine, and tooling for expressing, testing, and evaluating composable language rules on input strings
Facebook · 2017
Later among the works it cites.
Generation of large-scale simulated utterances in virtual rooms to train deep-neural networks for far-field speech recognition in google home
Chanwoo Kim, Ananya Misra, Kean Chin, Thad Hughes, Arun Narayanan, Tara Sainath, and Michiel Bacchiani · 2017
Later among the works it cites.
Just ask: Building an architecture for extensible self-service spoken language understanding
Anjishnu Kumar, Arpit Gupta, Julian Chan, Sam Tucker, Bjorn Hoffmeister, and Markus Dreyer · 2017
Later among the works it cites.
Backstitch: Counteracting finite-sample bias via negative steps
Yiming Wang, Vijayaditya Peddinti, Hainan Xu, Xiaohui Zhang, Daniel Povey, and Sanjeev Khudanpur · 2017
Later among the works it cites.
Low latency acoustic modeling using temporal convolution and lstms
Vijayaditya Peddinti, Yiming Wang, Daniel Povey, and Sanjeev Khudanpur · 2018
Closest in time.
Rustling, Rust implementation of Duckling
Snips Team · 2018
Closest in time.
Snips NLU rust, Snips NLU Rust implementation
Snips Team · 2018
Closest in time.
Snips NLU, Snips Python library to extract meaning from text
Snips Team · 2018
Closest in time.