Fetching the paper…
Reading the bibliography…
Scripted dialogues such as movie and TV subtitles constitute a widespread source of training data for conversational NLP models.
On getting a word in edgewise
Victor H. Yngve. 1970 · 1970
Earlier work this paper cites.
Side Sequences
Gail Jefferson. 1972 · 1972
Earlier work this paper cites.
Discourse as an interactional achievement: Some uses of ‘uh-huh’ and other things that come between sentences
Emanuel A. Schegloff. 1982 · 1981
Earlier work this paper cites.
Notes on a systematic deployment of the acknowledgement tokens “Yeah” and “MmHm”
Gail Jefferson. 1984 · 1984
Earlier work this paper cites.
Between and within: Alternative sequential treatments of continuers and assessments
Charles Goodwin. 1986 · 1986
Earlier work this paper cites.
Om det svenska systemet för språklig återkoppling
Jens Allwood. 1988 · 1988
Earlier work this paper cites.
Contributing to discourse
Herbert H. Clark and Edward F. Schaefer. 1989 · 1989
Earlier work this paper cites.
The HCRC map task corpus
Anne H. Anderson, Miles Bader, et al. 1991 · 1991
Earlier work this paper cites.
On the semantics and pragmatics of linguistic feedback
Jens Allwood, Joakim Nivre, and Elisabeth Ahlsén. 1992 · 1992
Earlier work this paper cites.
SWITCHBOARD: Telephone speech corpus for research and development
John J. Godfrey, Edward C. Holliman, and Jane McDaniel. 1992 · 1992
Earlier work this paper cites.
Context and dialogue control
Harry Bunt. 1994 · 1994
Earlier work this paper cites.
A Computational Theory of Grounding in Natural Language Conversation
David R. Traum. 1994 · 1994
Earlier work this paper cites.
Using Language
Herbert H. Clark. 1996 · 1996
Earlier work this paper cites.
Disfluencies in Switchboard
Elizabeth Shriberg. 1996 · 1996
Earlier work this paper cites.
CALLHOME Mandarin Chinese Transcripts
Barbara Wheatley. 1996 · 1996
Earlier work this paper cites.
Lexical, prosodic, and syntactic cues for dialog acts
Daniel Jurafsky, Elizabeth Shriberg, Barbara Fox, and Traci Curl. 1998 · 1998
Earlier work this paper cites.
Hollywood movie dialogue and the “real realism” of John Cassavetes
Todd Berliner. 1999 · 1999
Earlier work this paper cites.
CallHome Japanese corpus (in Japanese)
Yasuharu Den and John Fry. 2000 · 2000
Earlier work this paper cites.
Grounding criterion: Toward a formal theory of grounding
Tim Paek and Eric Horvitz. 2000 · 2000
Earlier work this paper cites.
Overlapping talk and the organization of turn-taking for conversation
Emanuel A. Schegloff. 2000 · 2000
Earlier work this paper cites.
An empirical study of acknowledgement structures
Philippe Muller and Laurent Prévot. 2003 · 2003
Earlier work this paper cites.
A Budapesti Szociolingvisztikai Interjú
Tamás Váradi. 2003 · 2003
Earlier work this paper cites.
The Fisher corpus: A resource for the next generations of speech-to-text
Christopher Cieri, David Miller, and Kevin Walker. 2004 · 2004
Earlier work this paper cites.
The Theory and Use of Clarification Requests in Dialogue
Matthew Purver. 2004 · 2004
Earlier work this paper cites.
Non-lexical conversational sounds in American English
Nigel Ward. 2006 · 2006
Cited alongside, same era.
The analysis of embodied communicative feedback in multimodal corpora: A prerequisite for behaviour simulation
Jens Allwood, Stefan Kopp, Karl Grammer, Elisabeh Ahlsén, Elisabeth Oberzaucher, and Markus Koppensteiner. 2007 · 2007
Cited alongside, same era.
Unleashing the killer corpus: experiences in creating the multi-everything AMI Meeting Corpus
Jean Carletta. 2007 · 2007
Cited alongside, same era.
Classifying Non-Sentential Utterances in Dialogue: A Machine Learning Approach
Raquel Fernández, Jonathan Ginzburg, and Shalom Lappin. 2007 · 2007
Cited alongside, same era.
An advanced speech corpus for Norwegian
Janne Bondi Johannessen, Kristin Hagen, Joel James Priestley, and Lars Nygaard. 2007 · 2007
Cited alongside, same era.
Prosodic features of very short utterances in dialogue
OpenSubtitles2016: Extracting large parallel corpora from movie and TV subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Later among the works it cites.
The ALICO corpus: Analysing the active listener
Zofia Malisz, Marcin Włodarczak, Hendrik Buschmeier, Joanna Skubisz, Stefan Kopp, and Petra Wagner. 2016 · 2016
Later among the works it cites.
A CUP of CoFee: A large collection of feedback utterances provided with communicative function annotations
Laurent Prévot, Jan Gorisch, and Roxane Bertrand. 2016 · 2016
Later among the works it cites.
The corpus of interactional data: A large multimodal annotated resource
Philippe Blache, Roxane Bertrand, Gaëlle Ferré, Berthille Pallaud, Laurent Prévot, and Stéphane Rauzy. 2017 · 2017
Later among the works it cites.
Using context information for dialog act classification in DNN framework
Yang Liu, Kun Han, Zhao Tan, and Yun Lei. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jens Edlund, Mattias Heldner, and Antoine Pelcé. 2009 · 2008
Cited alongside, same era.
Where do we backchannel?: On the use of mm, mhm, uh huh and such like
Göran Kjellmer. 2009 · 2009
Cited alongside, same era.
CLIPS diatopic, diamesic and diaphasic variations in spoken Italian
Renata Savy and Francesco Cutugno. 2009 · 2009
Cited alongside, same era.
Pitch similarity in the vicinity of backchannels
Mattias Heldner, Jens Edlund, and Julia Hirschberg. 2010 · 2010
Cited alongside, same era.
HAMATAC. The Hamburg MapTask Corpus
HZSK. 2010 · 2010
Cited alongside, same era.
Chameleons in imagined conversations: A new approach to understanding coordination of linguistic style in dialogs
Cristian Danescu-Niculescu-Mizil and Lillian Lee. 2011 · 2011
Cited alongside, same era.
A multimodal analysis of vocal and visual backchannels in spontaneous dialogs
Khiet P. Truong, Ronald Poppe, Iwan de Kok, and Dirk Heylen. 2011 · 2011
Cited alongside, same era.
Communicative listener feedback in human-agent interaction: Artificial speakers need to be attentive and adaptive
Hendrik Buschmeier and Stefan Kopp. 2018 · 2018
Later among the works it cites.
OpenSubtitles2018: Statistical rescoring of sentence alignments in large, noisy parallel corpora
Pierre Lison, Jörg Tiedemann, and Milen Kouylekov. 2018 · 2018
Later among the works it cites.
ISO-standard domain-independent dialogue act tagging for conversational agents
Stefano Mezza, Alessandra Cervone, Evgeny Stepanov, Giuliano Tortoreto, and Giuseppe Riccardi. 2018 · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Should we use movie subtitles to study linguistic patterns of conversational speech? A study based on French, English and Taiwan Mandarin
Laurent Prevot, Pierre Magistry, and Pierre Lison. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
What dialog is absent from constructed dialog?
Christoph Rühlemann. 2020 · 2020
Later among the works it cites.
The TV and movies corpora: Design, construction, and use
Mark Davies. 2021 · 2021
Later among the works it cites.
Feedback relevance spaces: Interactional constraints on processing contexts in Dynamic Syntax
Christine Howes and Arash Eshghi. 2021 · 2021
Later among the works it cites.
LoRA: Low-Rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Later among the works it cites.
Should we stop training more monolingual models, and simply use machine translation instead?
Tim Isbister, Fredrik Carlsson, and Magnus Sahlgren. 2021 · 2021
Later among the works it cites.
Large-scale text pre-training helps with dialogue act recognition, but not without fine-tuning
Bill Noble and Vladislav Maraev. 2021 · 2021
Later among the works it cites.
Modelling feedback in interaction with conversational agents – A review
Agnes Axelsson, Hendrik Buschmeier, and Gabriel Skantze. 2022 · 2022
Later among the works it cites.
Quantifying the interplay of conversational devices in building mutual understanding
Christina Dideriksen, Morten H Christiansen, Kristian Tylén, Mark Dingemanse, and Riccardo Fusaroli. 2022 · 2022
Later among the works it cites.
From text to talk: Harnessing conversational corpora for humane and diversity-aware language technology
Mark Dingemanse and Andreas Liesenfeld. 2022 · 2022
Later among the works it cites.
Bottom-up discovery of structure and variation in response tokens (‘backchannels’) across diverse languages
Andreas Liesenfeld and Mark Dingemanse. 2022 · 2022
Later among the works it cites.
Gemma: Open models based on gemini research and technology
Gemma Team, Thomas Mesnard, Cassidy Hardin, Robert Dadashi, Surya Bhupatiraju, Shreya Pathak, Laurent Sifre, Morgane Rivière, Mihir Sanjay Kale, Juliette Love, Pouya Tafti, Léonard Hussenot, Pier Giuseppe Sessa, Aakanksha Chowdhery, Adam Roberts, Aditya Barua, Alex Botev, Alex Castro-Ros, Ambrose Slone, Amélie Héliou, Andrea Tacchetti, Anna Bulanova, Antonia Paterson, Beth Tsai, Bobak Shahriari, Charline Le Lan, Christopher A Choquette-Choo, Clément Crepy, Daniel Cer, Daphne Ippolito, David Reid, Elena Buchatskaya, Eric Ni, Eric Noland, Geng Yan, George Tucker, George-Christian Muraru, Grigory Rozhdestvenskiy, Henryk Michalewski, Ian Tenney, Ivan Grishchenko, Jacob Austin, James Keeling, Jane Labanowski, Jean-Baptiste Lespiau, Jeff Stanway, Jenny Brennan, Jeremy Chen, Johan Ferret, Justin Chiu, Justin Mao-Jones, Katherine Lee, Kathy Yu, Katie Millican, Lars Lowe Sjoesund, Lisa Lee, Lucas Dixon, Machel Reid, Maciej Mikuła, Mateo Wirth, Michael Sharman, Nikolai Chinaev, Nithum Thain, Olivier Bachem, Oscar Chang, Oscar Wahltinez, Paige Bailey, Paul Michel, Petko Yotov, Rahma Chaabouni, Ramona Comanescu, Reena Jana, Rohan Anil, Ross McIlroy, Ruibo Liu, Ryan Mullins, Samuel L Smith, Sebastian Borgeaud, Sertan Girgin, Sholto Douglas, Shree Pandya, Siamak Shakeri, Soham De, Ted Klimenko, Tom Hennigan, Vlad Feinberg, Wojciech Stokowiec, Yu-Hui Chen, Zafarali Ahmed, Zhitao Gong, Tris Warkentin, Ludovic Peran, Minh Giang, Clément Farabet, Oriol Vinyals, Jeff Dean, Koray Kavukcuoglu, Demis Hassabis, Zoubin Ghahramani, Douglas Eck, Joelle Barral, Fernando Pereira, Eli Collins, Armand Joulin, Noah Fiedel, Evan Senter, Alek Andreev, and Kathleen Kenealy. 2024 · 2024
Closest in time.
Measures and mechanisms of common ground: Backchannels, conversational repair, and interactive alignment in free and task-oriented social interactions
Riccardo Fusaroli, Kristian Tylén, Katrine Garly, Jakob Steensig, Morten H. Christiansen, and Mark Dingemanse. 2017 · 2060
Closest in time.