Fetching the paper…
Reading the bibliography…
Recent advances in Large Language Models (LLMs) have highlighted the need for robust, comprehensive, and challenging benchmarks.
The equivalence of weighted kappa and the intraclass correlation coefficient as measures of reliability
Joseph L Fleiss and Jacob Cohen. 1973 · 1973
Earlier work this paper cites.
A psychoevolutionary theory of emotions
Robert Plutchik. 1982 · 1982
Earlier work this paper cites.
Expression and the nature of emotion
Paul Ekman. 1984 · 1984
Earlier work this paper cites.
Does the autistic child have a “theory of mind”?
Simon Baron-Cohen, Alan M Leslie, and Uta Frith. 1985 · 1985
Earlier work this paper cites.
Emotional intelligence
Peter Salovey and John D Mayer. 1990 · 1990
Earlier work this paper cites.
An advanced test of theory of mind: Understanding of story characters’ thoughts and feelings by able autistic, mentally handicapped, and normal children and adults
Francesca GE Happé. 1994 · 1994
Earlier work this paper cites.
Emotional intelligence. why it can matter more than iq
Daniel Goleman. 1996 · 1996
Earlier work this paper cites.
The media equation: How people treat computers, television, and new media like real people
Byron Reeves and Clifford Nass. 1996 · 1996
Earlier work this paper cites.
BarOn emotional quotient inventory , volume 40
Reuven Bar-On. 1997 · 1997
Earlier work this paper cites.
Recognition of faux pas by normally developing children and children with asperger syndrome or high-functioning autism
Simon Baron-Cohen, Michelle O’riordan, Valerie Stone, Rosie Jones, and Kate Plaisted. 1999 · 1999
Earlier work this paper cites.
Emotional intelligence meets traditional standards for an intelligence
John D Mayer, David R Caruso, and Peter Salovey. 1999 · 1999
Earlier work this paper cites.
Do autism spectrum disorders differ from each other and from non-spectrum disorders on emotion recognition tests?
Murray J Dyck, Kara Ferguson, and Ian M Shochet. 2001 · 2001
Earlier work this paper cites.
Toward machine emotional intelligence: Analysis of affective physiological state
Rosalind W. Picard, Elias Vyzas, and Jennifer Healey. 2001 · 2001
Earlier work this paper cites.
Emotional intelligence and interpersonal relations
Nicola S Schutte, John M Malouff, Chad Bobik, Tracie D Coston, Cyndy Greeson, Christina Jedlicka, Emily Rhodes, and Greta Wendorf. 2001 · 2001
Earlier work this paper cites.
Characteristic emotional intelligence and emotional well-being
Nicola S Schutte, John M Malouff, Maureen Simunek, Jamie McKenley, and Sharon Hollander. 2002 · 2002
Earlier work this paper cites.
Emotional intelligence and social interaction
Paulo N Lopes, Marc A Brackett, John B Nezlek, Astrid Schütz, Ina Sellin, and Peter Salovey. 2004 · 2004
Earlier work this paper cites.
Rumors of the death of emotional intelligence in organizational behavior are vastly exaggerated
Neal M Ashkanasy and Catherine S Daus. 2005 · 2005
Earlier work this paper cites.
A review and critique of emotional intelligence measures
Jeffrey M Conte. 2005 · 2005
Earlier work this paper cites.
Goemotions: A dataset of fine-grained emotions
Dorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko, Alan Cowen, Gaurav Nemade, and Sujith Ravi. 2020 · 2005
Earlier work this paper cites.
Multiple intelligences, the mozart effect, and emotional intelligence: A critical review
Lynn Waterhouse. 2006 · 2006
Earlier work this paper cites.
Mayer-salovey-caruso emotional intelligence test
John D Mayer, Peter Salovey, and David R Caruso. 2007 · 2007
Earlier work this paper cites.
New paradigms for assessing emotional intelligence: theory and data
Carolyn MacCann and Richard D Roberts. 2008 · 2008
Earlier work this paper cites.
Toward machines with emotional intelligence
Rosalind W Picard. 2008 · 2008
Cited alongside, same era.
Associations of trait and ability emotional intelligence with performance on theory of mind tasks in an adult sample
Fiona J Ferguson and Elizabeth J Austin. 2010 · 2010
Cited alongside, same era.
Dissociating cognitive from affective theory of mind: A tms study
Elke Kalbe, Marius Schlegel, Alexander T. Sack, Dennis A. Nowak, Manuel Dafotakis, Christopher Bangard, Matthias Brand, Simone Shamay-Tsoory, Oezguer A. Onur, and Josef Kessler. 2010 · 2010
Cited alongside, same era.
The involvement of emotion recognition in affective theory of mind
Daniela Mier, Stefanie Lis, Kerstin Neuthe, Carina Sauer, Christine Esslinger, Bernd Gallhofer, and Peter Kirsch. 2010 · 2010
Cited alongside, same era.
Emotional intelligence and agents: Survey and possible applications
Mirjana Ivanović, Miloš Radovanović, Zoran Budimac, Dejan Mitrović, Vladimir Kurbalija, Weihui Dai, and Weidong Zhao. 2014 · 2014
Cited alongside, same era.
CICERO: A dataset for contextualized commonsense inference in dialogues
Deepanway Ghosal, Siqi Shen, Navonil Majumder, Rada Mihalcea, and Soujanya Poria. 2022 · 2022
Later among the works it cites.
Sahand Sabour, Wen Zhang, Xiyao Xiao, Yuwei Zhang, Yinhe Zheng, Jiaxin Wen, Jialu Zhao, and Minlie Huang. 2022 · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Later among the works it cites.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Engagement, emotions, and relationships: on building intelligent agents
Candace L Sidner. 2016 · 2016
Cited alongside, same era.
Do we need emotionally intelligent artificial agents? first results of human perceptions of emotional intelligence in humans compared to robots
Lisa Fan, Matthias Scheutz, Monika Lohani, Marissa McCoy, and Charlene Stokes. 2017 · 2017
Cited alongside, same era.
Dailydialog: A manually labelled multi-turn dialogue dataset
Yanran Li, Hui Su, Xiaoyu Shen, Wenjie Li, Ziqiang Cao, and Shuzi Niu. 2017 · 2017
Cited alongside, same era.
Towards empathetic open-domain conversation models: A new benchmark and dataset
Hannah Rashkin, Eric Michael Smith, Margaret Li, and Y-Lan Boureau. 2018 · 2018
Cited alongside, same era.
The age of artificial emotional intelligence
Dagmar Schuller and Björn W Schuller. 2018 · 2018
Cited alongside, same era.
The measurement of emotional intelligence: A critical review of the literature and recommendations for researchers and practitioners
Peter J O’Connor, Andrew Hill, Maria Kaya, and Brett Martin. 2019 · 2019
Cited alongside, same era.
MELD: A multimodal multi-party dataset for emotion recognition in conversations
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, Gautam Naik, Erik Cambria, and Rada Mihalcea. 2019 · 2019
Cited alongside, same era.
Mostafa M. Amin, Rui Mao, Erik Cambria, and Björn W. Schuller. 2023 · 2023
Later among the works it cites.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, Kai Dang, Xiaodong Deng, Yang Fan, Wenbin Ge, Yu Han, Fei Huang, et al. 2023 · 2023
Later among the works it cites.
Hi-tom: A benchmark for evaluating higher-order theory of mind reasoning in large language models
Yinghui He, Yufan Wu, Yilin Jia, Rada Mihalcea, Yulong Chen, and Naihao Deng. 2023 · 2023
Later among the works it cites.
Who is chatgpt? benchmarking llms’ psychological portrayal using psychobench
Jen-tse Huang, Wenxuan Wang, Eric John Li, Man Ho Lam, Shujie Ren, Youliang Yuan, Wenxiang Jiao, Zhaopeng Tu, and Michael R Lyu. 2023 · 2023
Later among the works it cites.
CRoW: Benchmarking commonsense reasoning in real-world tasks
Mete Ismayilzada, Debjit Paul, Syrielle Montariol, Mor Geva, and Antoine Bosselut. 2023 · 2023
Later among the works it cites.
FANToM: A benchmark for stress-testing machine theory of mind in interactions
Hyunwoo Kim, Melanie Sclar, Xuhui Zhou, Ronan Bras, Gunhee Kim, Yejin Choi, and Maarten Sap. 2023 · 2023
Later among the works it cites.
Large language models understand and can be enhanced by emotional stimuli
Cheng Li, Jindong Wang, Yixuan Zhang, Kaijie Zhu, Wenxin Hou, Jianxun Lian, Fang Luo, Qiang Yang, and Xing Xie. 2023 · 2023
Later among the works it cites.
Tomchallenges: A principle-guided dataset and diverse evaluation tasks for exploring theory of mind
Xiaomeng Ma, Lingyu Gao, and Qihui Xu. 2023 · 2023
Later among the works it cites.
OpenAI. 2023 · 2023
Later among the works it cites.
Eq-bench: An emotional intelligence benchmark for large language models
Samuel J Paech. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Later among the works it cites.
Large language models fail on trivial alterations to theory-of-mind tasks
Tomer Ullman. 2023 · 2023
Later among the works it cites.
Emotional intelligence of large language models
Xuena Wang, Xueting Li, Zi Yin, Yue Wu, and Jia Liu. 2023 · 2023
Later among the works it cites.
Efficient cross-task prompt tuning for few-shot conversational emotion recognition
Yige Xu, Zhiwei Zeng, and Zhiqi Shen. 2023 · 2023
Later among the works it cites.
Safetybench: Evaluating the safety of large language models with multiple choice questions
Zhexin Zhang, Leqi Lei, Lindong Wu, Rui Sun, Yongkang Huang, Chong Long, Xiao Liu, Xuanyu Lei, Jie Tang, and Minlie Huang. 2023 · 2023
Later among the works it cites.
Large language models are not robust multiple choice selectors
Chujie Zheng, Hao Zhou, Fandong Meng, Jie Zhou, and Minlie Huang. 2023 · 2023
Later among the works it cites.
Agieval: A human-centric benchmark for evaluating foundation models
Wanjun Zhong, Ruixiang Cui, Yiduo Guo, Yaobo Liang, Shuai Lu, Yanlin Wang, Amin Saied, Weizhu Chen, and Nan Duan. 2023 · 2023
Later among the works it cites.