Fetching the paper…
Reading the bibliography…
We present an end-to-end approach to extract semantic concepts directly from the speech audio signal.
J. L. Elman, “Learning and development in neural networks: The importance of starting small,”
1993
Earlier work this paper cites.
A. L. Gorin, G. Riccardi, and J. H. Wright, “How may i help you?”
1997
Earlier work this paper cites.
F. Kubala, R. Schwartz, R. Stone, and R. Weischedel, “Named entity extraction from speech,” in
1998
Earlier work this paper cites.
H. Bonneau-Maynard, S. Rosset, C. Ayache, A. Kuhn, and D. Mostefa, “Semantic annotation of the French MEDIA dialog corpus,” in
2005
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
S. Yaman, L. Deng, D. Yu, Y.-Y. Wang, and A. Acero, “An integrative and discriminative technique for spoken utterance classification,”
2008
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in
2009
Earlier work this paper cites.
K. A. Krueger and P. Dayan, “Flexible shaping: How learning in small steps helps,”
2009
Earlier work this paper cites.
S. Galliano, G. Gravier, and L. Chaubard, “The ESTER 2 evaluation campaign for the rich transcription of French radio broadcasts,” in
2009
Earlier work this paper cites.
Y. Estève, T. Bazillon, J.-Y. Antoine, F. Béchet, and J. Farinas, “The EPAC corpus: Manual and automatic annotations of conversational speech in French broadcast news.” in
2010
Earlier work this paper cites.
T. Lavergne, O. Cappé, and F. Yvon, “Practical very large scale crfs,” in
2010
Earlier work this paper cites.
A. Nasr, F. Béchet, and J.-F. Rey, “Macaon: Une chaîne linguistique pour le traitement de graphes de mots,” in
2010
Earlier work this paper cites.
G. Tur and R. De Mori,
2011
Cited alongside, same era.
Y. Bengio, “Deep learning of representations for unsupervised and transfer learning,” in
2011
Cited alongside, same era.
S. Hahn, M. Dinarelli, C. Raymond, F. Lefevre, P. Lehnen, R. De Mori, A. Moschitti, H. Ney, and G. Riccardi, “Comparing stochastic approaches to spoken language understanding in multiple languages,”
2011
Cited alongside, same era.
C. Grouin, S. Rosset, P. Zweigenbaum, K. Fort, O. Galibert, and L. Quintard, “Proposal for an extension of traditional named entities: From guidelines to evaluation, an overview,” in
2011
Cited alongside, same era.
G. Tur, L. Deng, D. Hakkani-Tür, and X. He, “Towards deeper understanding: Deep convex networks for semantic utterance classification,” in
2012
Cited alongside, same era.
A. Bérard, O. Pietquin, L. Besacier, and C. Servan, “Listen and translate: A proof of concept for end-to-end speech-to-text translation,” in
2016
Later among the works it cites.
D. Amodei, S. Ananthanarayanan, R. Anubhai, J. Bai, E. Battenberg, C. Case, J. Casper, B. Catanzaro, Q. Cheng, G. Chen
2016
Later among the works it cites.
E. Simonnet, S. Ghannay, N. Camelin, Y. Estève, and R. De Mori, “ASR error management for improving spoken language understanding,” in
2017
Later among the works it cites.
R. J. Weiss, J. Chorowski, N. Jaitly, Y. Wu, and Z. Chen, “Sequence-to-sequence models can directly translate foreign speech,”
2017
Later among the works it cites.
S. Ghannay, A. Caubrière, Y. Estève, N. Camelin, E. Simonnet, A. Laurent, and E. Morin, “End-to-end named entity and semantic concept extraction from speech,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Lefèvre, D. Mostefa, L. Besacier, Y. Estève, M. Quignard, N. Camelin, B. Favre, B. Jabaian, and L. M. R. Barahona, “Leveraging study of robustness and portability of spoken language understanding systems across languages and domains: the PORTMEDIA corpora,” in
2012
Cited alongside, same era.
B. Jabaian, F. Lefèvre, and L. Besacier, “Portability of semantic annotations for fast development of dialogue corpora,” in
2012
Cited alongside, same era.
G. Gravier, G. Adda, N. Paulson, M. Carré, A. Giraudel, and O. Galibert, “The ETAPE corpus for the evaluation of speech-based TV content processing in the French language,” in
2012
Cited alongside, same era.
A. Giraudel, M. Carré, V. Mapelli, J. Kahn, O. Galibert, and L. Quintard, “The REPERE corpus: a multimodal corpus for person recognition.” in
2012
Cited alongside, same era.
M. Morchid, G. Linares, M. El-Beze, and R. De Mori, “Theme identification in telephone service conversations using quaternions of speech features.” in
2013
Cited alongside, same era.
Y.-N. Chen, W. Y. Wang, and A. I. Rudnicky, “Unsupervised induction and filling of semantic slots for spoken dialogue systems using frame-semantic parsing,” in
2013
Cited alongside, same era.
2014
Cited alongside, same era.
A. Bérard, L. Besacier, A. C. Kocabiyikoglu, and O. Pietquin, “End-to-end automatic speech translation of audiobooks,” in
2018
Later among the works it cites.
N. Jan, R. Cattoni, S. Sebastian, M. Cettolo, M. Turchi, and M. Federico, “The iwslt 2018 evaluation campaign,” in
2018
Later among the works it cites.
D. Serdyuk, Y. Wang, C. Fuegen, A. Kumar, B. Liu, and Y. Bengio, “Towards end-to-end spoken language understanding,” in
2018
Later among the works it cites.
2018
Later among the works it cites.
2019
Closest in time.
N. Tomashenko, A. Caubrière, and Y. Estève, “Investigating adaptation and transfer learning for end-to-end spoken language understanding from speech,” in
2019
Closest in time.