Fetching the paper…
Reading the bibliography…
In this paper, we present a database of emotional speech intended to be open-sourced and used for synthesis and generation purpose.
Ekman, P.: Basic emotions. In: Dalgleish, T., Powers, M.J. (eds.) Handbook of Cognition and Emotion, pp. 4–5. Wiley (1999)
1999
Earlier work this paper cites.
Trouvain, J.: Phonetic aspects of “speech-laughs”. In: Oralité et Gestualité: Actes du colloque ORAGE, Aix-en-Provence. Paris: L’Harmattan. pp. 634–639 (2001)
2001
Earlier work this paper cites.
Kawanami, H., Iwami, Y., Toda, T., Saruwatari, H., Shikano, K.: Gmm-based voice conversion applied to emotional speech synthesis. In: Eighth European Conference on Speech Communication and Technology (2003)
2003
Earlier work this paper cites.
Kominek, J., Black, A.W.: The cmu arctic speech databases. In: Fifth ISCA Workshop on Speech Synthesis (2004)
2004
Earlier work this paper cites.
Burkhardt, F., Paeschke, A., Rolfes, M., Sendlmeier, W.F., Weiss, B.: A database of german emotional speech. In: Ninth European Conference on Speech Communication and Technology (2005)
2005
Earlier work this paper cites.
Posner, J., Russell, J.A., Peterson, B.S.: The circumplex model of affect: an integrative approach to affective neuroscience, cognitive development, and psychopathology. Development and psychopathology 17 3
2005
Earlier work this paper cites.
Busso, C., Bulut, M., Lee, C.C., Kazemzadeh, A., Mower, E., Kim, S., Chang, J.N., Lee, S., Narayanan, S.S.: Iemocap: Interactive emotional dyadic motion capture database. Language resources and evaluation 42
2008
Earlier work this paper cites.
Bänziger, T., Mortillaro, M., Scherer, K.R.: Introducing the geneva multimodal expression corpus for experimental research on emotion perception. Emotion 12
2012
Earlier work this paper cites.
Brognaux, S., Roekhaut, S., Drugman, T., Beaufort, R.: Train&Align: A new online tool for automatic phonetic alignment. In: Spoken Language Technology Workshop (SLT), 2012 IEEE. pp. 416–421. IEEE (2012)
2012
Earlier work this paper cites.
Cao, H., Cooper, D.G., Keutmann, M.K., Gur, R.C., Nenkova, A., Verma, R.: Crema-d: Crowd-sourced emotional multimodal actors dataset. IEEE transactions on affective computing 5
2014
Cited alongside, same era.
El Haddad, K., Cakmak, H., Dupont, S., , Dutoit, T.: Breath and repeat: An attempt at enhancing speech-laugh synthesis quality. In: European Signal Processing Conference (EUSIPCO 2015). Nice, France (31 August-4 September 2015)
2015
Cited alongside, same era.
El Haddad, K., Cakmak, H., Dupont, S., Dutoit, T.: An HMM Approach for Synthesizing Amused Speech with a Controllable Intensity of Smile. In: IEEE International Symposium on Signal Processing and Information Technology (ISSPIT). Abu Dhabi, UAE (7-10 December 2015)
2015
Cited alongside, same era.
El Haddad, K., Dupont, S., d’Alessandro, N., Dutoit, T.: An HMM-based speech-smile synthesis system: An approach for amusement synthesis. In: International Workshop on Emotion Representation, Analysis and Synthesis in Continuous Time and Space (EmoSPACE). Ljubljana, Slovenia (4-8 May 2015)
Wu, Z., Watts, O., King, S.: Merlin: An open source neural network speech synthesis system. Proc. SSW, Sunnyvale, USA (2016)
2016
Later among the works it cites.
Busso, C., Parthasarathy, S., Burmania, A., AbdelWahab, M., Sadoughi, N., Provost, E.M.: Msp-improv: An acted corpus of dyadic interactions to study emotion perception. IEEE Transactions on Affective Computing 8
2017
Later among the works it cites.
El Haddad, K., Torre, I., Gilmartin, E., Çakmak, H., Dupont, S., Dutoit, T., Campbell, N.: Introducing amus: The amused speech database. In: Camelin, N., Estève, Y., Martín-Vide, C. (eds.) Statistical Language and Speech Processing. pp. 229–240. Springer International Publishing, Cham (2017)
2017
Later among the works it cites.
Honnet, P.E., Lazaridis, A., Garner, P.N., Yamagishi, J.: The siwis french speech synthesis database ? design and recording of a high quality french database for speech synthesis. Online Database (2017)
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
El Haddad, K., Dupont, S., Urbain, J., Dutoit, T.: Speech-laughs: An HMM-based Approach for Amused Speech Synthesis. In: Internation Conference on Acoustics, Speech and Signal Processing (ICASSP 2015). pp. 4939–4943. Brisbane, Australia (19-24 April 2015)
2015
Cited alongside, same era.
Luo, Z., Takiguchi, T., Ariki, Y.: Emotional voice conversion using deep neural networks with mcc and f0 features. In: Computer and Information Science (ICIS), 2016 IEEE/ACIS 15th International Conference on. pp. 1–5. IEEE (2016)
2016
Cited alongside, same era.
Morise, M., Yokomori, F., Ozawa, K.: World: a vocoder-based high-quality speech synthesis system for real-time applications. IEICE TRANSACTIONS on Information and Systems 99
2016
Cited alongside, same era.
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A.W., Kavukcuoglu, K.: Wavenet: A generative model for raw audio. In: SSW (2016)
2016
Cited alongside, same era.
2017
Later among the works it cites.
Wang, Y., Skerry-Ryan, R.J., Stanton, D., Wu, Y., Weiss, R.J., Jaitly, N., Yang, Z., Xiao, Y., Chen, Z., Bengio, S., Le, Q.V., Agiomyrgiannakis, Y., Clark, R., Saurous, R.A.: Tacotron: Towards end-to-end speech synthesis. In: INTERSPEECH (2017)
2017
Later among the works it cites.
Yannakakis, G.N., Cowie, R., Busso, C.: The ordinal nature of emotions. In: 2017 Seventh International Conference on Affective Computing and Intelligent Interaction (ACII). vol. 00, pp. 248–255 (Oct 2017). https://doi.org/10.1109/ACII.2017.8273608, doi.ieeecomputersociety.org/10.1109/ACII.2017.8273608
2017
Later among the works it cites.
Livingstone, S.R., Russo, F.A.: The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english. PLOS ONE 13
2018
Closest in time.