Fetching the paper…
Reading the bibliography…
Imbuing machines with the ability to talk has been a longtime pursuit of artificial intelligence (AI) research.
“Homer. The Iliad. With an English translation by AT Murray. Vol. I. Loeb Classical Library. London: William Heinemann. New York: GP Putman’s Sons, 1924. Pp. Xviii+ 579”
Samuel Bassett · 1925
Earlier work this paper cites.
“An artificial talker driven from a phonetic input”
John Kelly and Louis Gerstman · 1961
Earlier work this paper cites.
“How to do things with words”
John Austin · 1975
Earlier work this paper cites.
“A model of articulatory dynamics and control”
Cecil Coker · 1976
Earlier work this paper cites.
“MITalk-79: The 1979 MIT text-to-speech system”
Jonathan Allen, Sharon Hunnicutt, Rolf Carlson and Bjorn Granstrom · 1979
Earlier work this paper cites.
“Vocal affect expression: a review and a model for future research.”
Klaus Scherer · 1986
Earlier work this paper cites.
“Review of text-to-speech conversion for English”
Dennis Klatt · 1987
Earlier work this paper cites.
“Generating expression in synthesized speech”, 1989
Janet Cahn · 1989
Earlier work this paper cites.
“Simulating emotion in synthetic speech”, 1989
Iain Murray · 1989
Earlier work this paper cites.
“The generation of affect in synthesized speech”
Janet Cahn · 1990
Earlier work this paper cites.
“Pitch-synchronous waveform processing techniques for text-to-speech synthesis using diphones”
Eric Moulines and Francis Charpentier · 1990
Earlier work this paper cites.
“An argument for basic emotions”
Paul Ekman · 1992
Earlier work this paper cites.
“Toward the simulation of emotion in synthetic speech: A review of the literature on human vocal emotion”
Iain Murray and John Arnott · 1993
Earlier work this paper cites.
“Conceptual pacts and lexical choice in conversation.”
Susan Brennan and Herbert Clark · 1996
Earlier work this paper cites.
“Progress in speech synthesis”
Jan Van, Richard Sproat, Joseph Olive and Julia Hirschberg · 1997
Earlier work this paper cites.
“Stress and emotion: A new synthesis”
Richard Lazarus · 1999
Earlier work this paper cites.
“Verification of acoustical correlates of emotional speech using formant synthesis”
Felix Burkhardt and Walter Sendlmeier · 2000
Earlier work this paper cites.
“Gibbs sampling”
Alan Gelfand · 2000
Earlier work this paper cites.
“Speech parameter generation algorithms for HMM-based speech synthesis”
Keiichi Tokuda et al · 2000
Earlier work this paper cites.
“The attention economy”
Thomas Davenport and John Beck · 2001
Earlier work this paper cites.
“Emotional speech synthesis: A review”
M. Schr“”oder · 2001
Earlier work this paper cites.
“Creating friendly AI 1.0: The analysis and design of benevolent goal architectures”
Eliezer Yudkowsky · 2001
Earlier work this paper cites.
“Training products of experts by minimizing contrastive divergence”
Geoffrey Hinton · 2002
Earlier work this paper cites.
“Unit selection and emotional speech”
Alan Black · 2003
Earlier work this paper cites.
“Describing the emotional states that are expressed in speech”
Roddy Cowie and Randolph Cornelius · 2003
Earlier work this paper cites.
“A corpus-based speech synthesis system with emotion”
Akemi Iida, Nick Campbell, Fumito Higuchi and Michiaki Yasumura · 2003
Earlier work this paper cites.
“Why Are Commercials so Loud? Perception and Modeling of the Loudness of Amplitude-Compressed Speech”
Brian Moore, Brian Glasberg and Michael Stone · 2003
Earlier work this paper cites.
“Vocal communication of emotion: A review of research paradigms”
Klaus Scherer · 2003
Earlier work this paper cites.
“Assessing dimensions of perceived visual aesthetics of web sites”
Talia Lavie and Noam Tractinsky · 2004
Earlier work this paper cites.
“Conversational agents”
James Lester, Karl Branting and Bradford Mott · 2004
Earlier work this paper cites.
“HMM-based speech synthesis with various speaking styles using model interpolation”
Makoto Tachibana et al · 2004
Earlier work this paper cites.
“Language, gender, and sexuality: Current issues and new directions”
Deborah Cameron · 2005
Earlier work this paper cites.
“Estimation of non-normalized statistical models by score matching.”
Aapo Hyv“”arinen and Peter Dayan · 2005
Earlier work this paper cites.
“Prosody conversion from neutral speech to emotional speech”
Jianhua Tao, Yongguo Kang and Aijun Li · 2006
Earlier work this paper cites.
“The voice conveys specific emotions: evidence from vocal burst displays.”
Emiliana Simon-Thomas et al · 2009
Earlier work this paper cites.
“Voice transformation: a survey”
Yannis Stylianou · 2009
Earlier work this paper cites.
“Statistical parametric speech synthesis”
Heiga Zen, Keiichi Tokuda and Alan Black · 2009
Cited alongside, same era.
“Comparison of emotion perception among different cultures”
Jianwu Dang et al · 2010
Cited alongside, same era.
“Compassion: an evolutionary analysis and empirical review.”
Jennifer Goetz, Dacher Keltner and Emiliana Simon-Thomas · 2010
Cited alongside, same era.
“Christian Gottlieb Kratzenstein: Pioneer in Speech Synthesis”
John Ohala · 2011
Cited alongside, same era.
“Policy shaping: Integrating human feedback with reinforcement learning”
Shane Griffith et al · 2013
Cited alongside, same era.
“Mediatization and sociolinguistic change. Key concepts, research traditions, open issues”
Jannis Androutsopoulos · 2014
Cited alongside, same era.
“How the voice persuades”
Alex Van and Jonah Berger · 2020
Later among the works it cites.
“On the opportunities and risks of foundation models”
Rishi Bommasani et al · 2021
Later among the works it cites.
“Semantic space theory: A computational approach to emotion”
Alan Cowen and Dacher Keltner · 2021
Later among the works it cites.
“Grad-TTS: A diffusion probabilistic model for text-to-speech”
Vadim Popov et al · 2021
Later among the works it cites.
“A survey on neural speech synthesis”
Xu Tan, Tao Qin, Frank Soong and Tie-Yan Liu · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“How do people compare themselves with others on social network sites? The case of Facebook”
Sang Lee · 2014
Cited alongside, same era.
“Computational Paralinguistics: Emotion, Affect and Personality in Speech and Language Processing”
Bj“”orn Schuller and Anton Batliner · 2014
Cited alongside, same era.
“Social comparison, social media, and self-esteem.”
Erin Vogel, Jason Rose, Lindsay Roberts and Katheryn Eckles · 2014
Cited alongside, same era.
“Advanced social engineering attacks”
Katharina Krombholz, Heidelinde Hobel, Markus Huber and Edgar Weippl · 2015
Cited alongside, same era.
“Tutorial on variational autoencoders”
Carl Doersch · 2016
Cited alongside, same era.
“Talk “like a man”: The linguistic styles of Hillary Clinton, 1992–2013”
Jennifer Jones · 2016
Cited alongside, same era.
Alice Baird et al · 2022
Later among the works it cites.
“Alexa as an active listener: how backchanneling can elicit self-disclosure and promote user experience”
Eugene Cho, Nasim Motalebi, S Sundar and Saeed Abdullah · 2022
Later among the works it cites.
“The acoustically emotion-aware conversational agent with speech emotion recognition and empathetic responses”
Jiaxiong Hu, Yun Huang, Xiaozhu Hu and Yingqing Xu · 2022
Later among the works it cites.
“Prodiff: Progressive fast diffusion model for high-quality text-to-speech”
Rongjie Huang et al · 2022
Later among the works it cites.
“Training Socially Engaging Robots: Modeling Backchannel Behaviors with Batch Reinforcement Learning”
Nusrah Hussain, Engin Erzin, T Sezgin and Yucel Yemez · 2022
Later among the works it cites.
“Metaverse”
Stylianos Mystakidis · 2022
Later among the works it cites.
“False idols: Unpacking the opportunities and challenges of falsity in the context of virtual influencers”
Sean Sands, Carla Ferraro, Vlad Demsar and Garreth Chandler · 2022
Later among the works it cites.
“Emergent Abilities of Large Language Models”
Jason Wei et al · 2022
Later among the works it cites.
“Emotion Intensity and its Control for Emotional Voice Conversion”
Kun Zhou et al · 2022
Later among the works it cites.
“Speech synthesis with mixed emotions”
Kun Zhou et al · 2022
Later among the works it cites.
Josh Achiam et al · 2023
Later among the works it cites.
“TikTok Voice Effects” Accessed: 22/04/2024, 2023
Alec Chillingworth · 2023
Later among the works it cites.
“Generative AI and ChatGPT: Applications, challenges, and AI-human collaboration”
Fiona Fui-Hoon et al · 2023
Later among the works it cites.
“PromptTTS: Controllable text-to-speech with text descriptions”
Zhifang Guo et al · 2023
Later among the works it cites.
“Evaluating the persuasive influence of political microtargeting with large language models”
Kobi Hackenburg and Helen Margetts · 2023
Later among the works it cites.
“Generative artificial intelligence in marketing: Applications, opportunities, challenges, and research agenda”
Nir Kshetri, Yogesh Dwivedi, Thomas Davenport and Niki Panteli · 2023
Later among the works it cites.
“ASVSpoof 2021: Towards spoofed and deepfake speech detection in the wild”
Xuechen Liu et al · 2023
Later among the works it cites.
“Consistency models”
Yang Song, Prafulla Dhariwal, Mark Chen and Ilya Sutskever · 2023
Later among the works it cites.
“Amendments adopted by the European Parliament on 14 June 2023 on the proposal for a regulation of the European Parliament and of the Council on laying down harmonised rules on artificial intelligence (Artificial Intelligence Act) and amending certain Union legislative acts (COM(2021)0206 – C9-0146/2021 – 2021/0106(COD))” https://www.europarl.europa.eu/doceo/document/TA-9-2023-0236_EN.html , 2023
The European Parliament · 2023
Later among the works it cites.
“Llama 2: Open foundation and fine-tuned chat models”
Hugo Touvron et al · 2023
Later among the works it cites.
“An Overview of Affective Speech Synthesis and Conversion in the Deep Learning Era”
Andreas Triantafyllopoulos et al · 2023
Later among the works it cites.
“Uniaudio: An audio foundation model toward universal audio generation”
Dongchao Yang et al · 2023
Later among the works it cites.
“Diffusion models: A comprehensive survey of methods and applications”
Ling Yang et al · 2023
Later among the works it cites.
“Deep learning-based expressive speech synthesis: a systematic review of approaches, challenges, and resources”
Huda Barakat, Oytun Turk and Cenk Demiroglu · 2024
Closest in time.
“A survey on generative diffusion models”
Hanqun Cao et al · 2024
Closest in time.
“Measuring the Persuasiveness of Language Models”, 2024
Esin Durmus et al · 2024
Closest in time.
“The benefits, risks and bounds of personalizing the alignment of large language models to individuals”
Hannah Kirk, Kertie Vidgen, Paul Röttger and Scott. Hale · 2024
Closest in time.
“PromptTTS 2: Describing and generating voices with text prompt”
Yichong Leng et al · 2024
Closest in time.
“EMOCONV-Diff: Diffusion-Based Speech Emotion Conversion for Non-Parallel and in-the-Wild Data”
Navin Prabhu et al · 2024
Closest in time.
“On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial”
Francesco Salvi, Manoel Ribeiro, Riccardo Gallotti and Robert West · 2024
Closest in time.
“PromptTTS++: Controlling Speaker Identity in Prompt-Based Text-to-Speech Using Natural Language Descriptions”
Reo Shimizu et al · 2024
Closest in time.