Fetching the paper…
Reading the bibliography…
Addressing the critical shortage of mental health resources for effective screening, diagnosis, and treatment remains a significant challenge.
Prosodic features which cue back-channel responses in english and japanese,
N. Ward, W. Tsukahara, · 2000
Earlier work this paper cites.
The prosody of backchannels in american english (2007)
S. Benus, A. Gravano, J. B. Hirschberg, · 2007
Earlier work this paper cites.
Backchannels revisited from a multimodal perspective,
R. Bertrand, G. Ferré, P. Blache, R. Espesser, S. Rauzy, · 2007
Earlier work this paper cites.
All smiles are not created equal: Morphology and timing of smiles perceived as amused, polite, and embarrassed/nervous,
Z. Ambadar, J. F. Cohn, L. I. Reed, · 2009
Earlier work this paper cites.
Backchannel-inviting cues in task-oriented dialogue,
A. Gravano, J. Hirschberg, · 2009
Earlier work this paper cites.
Opensmile: the munich versatile and fast open-source audio feature extractor,
F. Eyben, M. Wöllmer, B. Schuller, · 2010
Earlier work this paper cites.
A multimodal analysis of vocal and visual backchannels in spontaneous dialogs.,
K. P. Truong, R. Poppe, I. de Kok, D. Heylen, · 2011
Earlier work this paper cites.
Detecting depression severity from vocal prosody,
Y. Yang, C. Fairbairn, J. F. Cohn, · 2012
Earlier work this paper cites.
Furhat: a back-projected human-like robot head for multiparty human-machine interaction,
S. Al Moubayed, J. Beskow, G. Skantze, B. Granström, · 2012
Earlier work this paper cites.
Simsensei kiosk: A virtual human interviewer for healthcare decision support,
D. DeVault, R. Artstein, G. Benn, T. Dey, E. Fast, A. Gainer, K. Georgila, J. Gratch, A. Hartholt, M. Lhommet, et al., · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate,
D. Bahdanau, K. Cho, Y. Bengio, · 2014
Cited alongside, same era.
J. W. Pennebaker, R. L. Boyd, K. Jordan, K. Blackburn, The development and psychometric properties of LIWC2015, Technical Report, 2015
2015
Cited alongside, same era.
Learn2smile: Learning non-verbal interaction through observation,
W. Feng, A. Kannan, G. Gkioxari, C. L. Zitnick, · 2017
Cited alongside, same era.
Montreal Forced Aligner: Trainable Text-Speech Alignment Using Kaldi,
M. McAuliffe, M. Socolof, S. Mihuc, M. Wagner, M. Sonderegger, · 2017
Cited alongside, same era.
Cnn architectures for large-scale audio classification,
S. Hershey, S. Chaudhuri, D. P. Ellis, J. F. Gemmeke, A. Jansen, R. C. Moore, M. Plakal, D. Platt, R. A. Saurous, B. Seybold, et al., · 2017
Cited alongside, same era.
wav2vec: Unsupervised pre-training for speech recognition,
S. Schneider, A. Baevski, R. Collobert, M. Auli, · 2019
Later among the works it cites.
Spectral representation of behaviour primitives for depression analysis,
S. Song, S. Jaiswal, L. Shen, M. Valstar, · 2020
Later among the works it cites.
Acoustic correlates of the voice qualifiers: A survey,
S. A. Memon, · 2020
Later among the works it cites.
No gestures left behind: Learning relationships between spoken language and freeform gestures,
C. Ahuja, D. W. Lee, R. Ishii, L.-P. Morency, · 2020
Later among the works it cites.
Exploring barriers to mental health care in the u.s. (2022). doi: 10.15766/rai_a3ewcf9p
H. Modi, K. Orgera, A. Grover, · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Every smile is unique: Landmark-guided diverse smile generation,
W. Wang, X. Alameda-Pineda, D. Xu, P. Fua, E. Ricci, N. Sebe, · 2018
Cited alongside, same era.
Sign language production using neural machine translation and generative adversarial networks,
S. Stoll, N. C. Camgöz, S. Hadfield, R. Bowden, · 2018
Cited alongside, same era.
Collaborative user responses in multiparty interaction with a couples counselor robot,
D. Utami, T. Bickmore, · 2019
Cited alongside, same era.
Afar: A deep learning based tool for automated facial affect recognition,
I. O. Ertugrul, L. A. Jeni, W. Ding, J. F. Cohn, · 2019
Cited alongside, same era.
Multimodal temporal machine learning for bipolar disorder and depression recognition,
F. Ceccarelli, M. Mahmoud, · 2022
Later among the works it cites.
Learning to listen: Modeling non-deterministic dyadic facial motion,
E. Ng, H. Joo, L. Hu, H. Li, T. Darrell, A. Kanazawa, S. Ginosar, · 2022
Later among the works it cites.
Voice activity projection: Self-supervised learning of turn-taking events,
E. Ekstedt, G. Skantze, · 2022
Later among the works it cites.
Affective faces for goal-driven dyadic communication,
S. Geng, R. Teotia, P. Tendulkar, S. Menon, C. Vondrick, · 2023
Later among the works it cites.