Fetching the paper…
Reading the bibliography…
In recent years, with the rapid advancements in large language models (LLMs), achieving excellent empathetic response capability has become a crucial prerequisite.
Measuring individual differences in empathy: Evidence for a multidimensional approach
Mark H Davis. 1983 · 1983
Earlier work this paper cites.
Adaptive mixtures of local experts
Robert A Jacobs, Michael I Jordan, Steven J Nowlan, and Geoffrey E Hinton. 1991 · 1991
Earlier work this paper cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz, Andy Davis, Quoc Le, Geoffrey Hinton, and Jeff Dean. 2017 · 2017
Earlier work this paper cites.
COMET: Commonsense Transformers for Automatic Knowledge Graph Construction. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . 4762–4779
Antoine Bosselut, Hannah Rashkin, Maarten Sap, Chaitanya Malaviya, Asli Celikyilmaz, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
MoEL: Mixture of Empathetic Listeners. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . 121–132
Zhaojiang Lin, Andrea Madotto, Jamin Shin, Peng Xu, and Pascale Fung. 2019 · 2019
Earlier work this paper cites.
Towards Empathetic Open-domain Conversation Models: A New Benchmark and Dataset. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . 5370–5381
Hannah Rashkin, Eric Michael Smith, Margaret Li, and Y-Lan Boureau. 2019 · 2019
Earlier work this paper cites.
COSMIC: COmmonSense knowledge for eMotion Identification in Conversations. In Findings of the Association for Computational Linguistics: EMNLP 2020 . 2470–2481
Deepanway Ghosal, Navonil Majumder, Alexander Gelbukh, Rada Mihalcea, and Soujanya Poria. 2020 · 2020
Earlier work this paper cites.
DEBERTA: DECODING-ENHANCED BERT WITH DISENTANGLED ATTENTION. In International Conference on Learning Representations
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2020 · 2020
Earlier work this paper cites.
Gshard: Scaling giant models with conditional computation and automatic sharding
Dmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen, Orhan Firat, Yanping Huang, Maxim Krikun, Noam Shazeer, and Zhifeng Chen. 2020 · 2020
Earlier work this paper cites.
EmpDG: Multi-resolution Interactive Empathetic Dialogue Generation. In Proceedings of the 28th International Conference on Computational Linguistics . 4454–4466
Qintong Li, Hongshen Chen, Zhaochun Ren, Pengjie Ren, Zhaopeng Tu, and Zhumin Chen. 2020 · 2020
Earlier work this paper cites.
Pei Zhou, Pegah Jandaghi, Bill Yuchen Lin, Justin Cho, Jay Pujara, and Xiang Ren. 2021 · 2021
Earlier work this paper cites.
Wish I Can Feel What You Feel: A Neural Approach for Empathetic Response Generation. In Findings of the Association for Computational Linguistics: EMNLP 2022 . 922–933
Yangbin Chen and Chunfeng Liang. 2022 · 2022
Earlier work this paper cites.
Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity
William Fedus, Barret Zoph, and Noam Shazeer. 2022 · 2022
Earlier work this paper cites.
Emp-RFT: Empathetic Response Generation via Recognizing Feature Transitions between Utterances. In Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 4118–4128
Wongyu Kim, Youbin Ahn, Donghyun Kim, and Kyong-Ho Lee. 2022 · 2022
Earlier work this paper cites.
Branch-train-merge: Embarrassingly parallel training of expert language models
Margaret Li, Suchin Gururangan, Tim Dettmers, Mike Lewis, Tim Althoff, Noah A Smith, and Luke Zettlemoyer. 2022a · 2022
Earlier work this paper cites.
Cem: Commonsense-aware empathetic response generation. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 36. 11229–11237
Sahand Sabour, Chujie Zheng, and Minlie Huang. 2022 · 2022
Earlier work this paper cites.
Empathetic Dialogue Generation via Sensitive Emotion Recognition and Sensible Knowledge Selection. In Findings of the Association for Computational Linguistics: EMNLP 2022 . 4634–4645
Lanrui Wang, Jiangnan Li, Zheng Lin, Fandong Meng, Chenxu Yang, Weiping Wang, and Jie Zhou. 2022 · 2022
Earlier work this paper cites.
Improving Empathetic Dialogue Generation by Dynamically Infusing Commonsense Knowledge. In Findings of the Association for Computational Linguistics: ACL 2023 . 7858–7873
Hua Cai, Xuli Shen, Qing Xu, Weilin Shen, Xiaomei Wang, Weifeng Ge, Xiaoqing Zheng, and Xiangyang Xue. 2023 · 2023
Earlier work this paper cites.
Alpagasus: Training a better alpaca with fewer data
Lichang Chen, Shiyang Li, Jun Yan, Hai Wang, Kalpa Gunaratna, Vikas Yadav, Zheng Tang, Vijay Srinivasan, Tianyi Zhou, Heng Huang, et al · 2023
Cited alongside, same era.
Lingua manga: A generic large language model centric system for data curation
Zui Chen, Lei Cao, and Sam Madden. 2023a · 2023
Cited alongside, same era.
Mods: Model-oriented data selection for instruction tuning
Qianlong Du, Chengqing Zong, and Jiajun Zhang. 2023 · 2023
Cited alongside, same era.
How large language models will disrupt data management
Raul Castro Fernandez, Aaron J Elmore, Michael J Franklin, Sanjay Krishnan, and Chenhao Tan. 2023 · 2023
Cited alongside, same era.
E-CORE: Emotion Correlation Enhanced Empathetic Dialogue Generation. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing . 10568–10586
Enhancing Empathetic and Emotion Support Dialogue Generation with Prophetic Commonsense Inference
Lanrui Wang, Jiangnan Li, Chenxu Yang, Zheng Lin, and Weiping Wang. 2023 · 2023
Later among the works it cites.
Rethinking the Instruction Quality: LIFT is What You Need
Yang Xu, Yongqiang Yao, Yufan Huang, Mengnan Qi, Maoquan Wang, Bin Gu, and Neel Sundaresan. 2023 · 2023
Later among the works it cites.
Exploiting Emotion-Semantic Correlations for Empathetic Response Generation. In Findings of the Association for Computational Linguistics: EMNLP 2023 . 4826–4837
Zhou Yang, Zhaochun Ren, Wang Yufeng, Xiaofei Zhu, Zhihao Chen, Tiecheng Cai, Wu Yunbing, Yisong Su, Sibo Ju, and Xiangwen Liao. 2023 · 2023
Later among the works it cites.
Don’t Lose Yourself! Empathetic Response Generation via Explicit Self-Other Awareness. In Findings of the Association for Computational Linguistics: ACL 2023 . 13331–13344
Weixiang Zhao, Yanyan Zhao, Xin Lu, and Bing Qin. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fengyi Fu, Lei Zhang, Quan Wang, and Zhendong Mao. 2023 · 2023
Cited alongside, same era.
Megablocks: Efficient sparse training with mixture-of-experts
Trevor Gale, Deepak Narayanan, Cliff Young, and Matei Zaharia. 2023 · 2023
Cited alongside, same era.
CAB: Empathetic Dialogue Generation with Cognition, Affection and Behavior. In Database Systems for Advanced Applications: 28th International Conference, DASFAA 2023 . 597–606
Pan Gao, Donghong Han, Rui Zhou, Xuejiao Zhang, and Zikun Wang. 2023 · 2023
Cited alongside, same era.
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning. In The Twelfth International Conference on Learning Representations
Wei Liu, Weihao Zeng, Keqing He, Yong Jiang, and Junxian He. 2023 · 2023
Cited alongside, same era.
# InsTag: Instruction Tagging for Analyzing Supervised Fine-tuning of Large Language Models. In The Twelfth International Conference on Learning Representations
Keming Lu, Hongyi Yuan, Zheng Yuan, Runji Lin, Junyang Lin, Chuanqi Tan, Chang Zhou, and Jingren Zhou. 2023 · 2023
Cited alongside, same era.
Flexmoe: Scaling large-scale sparse pre-trained model training via dynamic device placement
Xiaonan Nie, Xupeng Miao, Zilong Wang, Zichao Yang, Jilong Xue, Lingxiao Ma, Gang Cao, and Bin Cui. 2023 · 2023
Cited alongside, same era.
GPT-4 technical report
R OpenAI. 2023b · 2023
Cited alongside, same era.
Think Twice: A Human-like Two-Stage Conversational Agent for Emotional Response Generation. In Proceedings of the 2023 International Conference on Autonomous Agents and Multiagent Systems . 727–736
Yushan Qian, Bo Wang, Shangzhao Ma, Wu Bin, Shuo Zhang, Dongming Zhao, Kun Huang, and Yuexian Hou. 2023a · 2023
Cited alongside, same era.
CASE: Aligning Coarse-to-Fine Cognition and Affection for Empathetic Response Generation. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics . 8223–8237
Jinfeng Zhou, Chujie Zheng, Bo Wang, Zheng Zhang, and Minlie Huang. 2023 · 2023
Later among the works it cites.
A Survey of Multimodal Large Language Model from A Data-centric Perspective
Tianyi Bai, Hao Liang, Binwang Wan, Ling Yang, Bozhou Li, Yifan Wang, Bin Cui, Conghui He, Binhang Yuan, and Wentao Zhang. 2024 · 2024
Closest in time.
Deepseekmoe: Towards ultimate expert specialization in mixture-of-experts language models
Damai Dai, Chengqi Deng, Chenggang Zhao, RX Xu, Huazuo Gao, Deli Chen, Jiashi Li, Wangding Zeng, Xingkai Yu, Y Wu, et al · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al · 2024
Closest in time.
Introducing Meta Llama 3: The most capable openly available LLM to date
meta llama. 2024 · 2024
Closest in time.
Demystifying Data Management for Large Language Models. In Companion of the 2024 International Conference on Management of Data . 547–555
Xupeng Miao, Zhihao Jia, and Bin Cui. 2024 · 2024
Closest in time.
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
Bowen Pan, Yikang Shen, Haokun Liu, Mayank Mishra, Gaoyuan Zhang, Aude Oliva, Colin Raffel, and Rameswar Panda. 2024 · 2024
Closest in time.
SelectLLM: Can LLMs Select Important Instructions to Annotate?
Ritik Sachin Parkar, Jaehyung Kim, Jong Inn Park, and Dongyeop Kang. 2024 · 2024
Closest in time.
Branch-Train-MiX: Mixing Expert LLMs into a Mixture-of-Experts LLM
Sainbayar Sukhbaatar, Olga Golovneva, Vasu Sharma, Hu Xu, Xi Victoria Lin, Baptiste Rozière, Jacob Kahn, Daniel Li, Wen-tau Yih, Jason Weston, et al · 2024
Closest in time.
Introducing Qwen1.5
Qwen Team. 2024 · 2024
Closest in time.
An Iterative Associative Memory Model for Empathetic Response Generation
Zhou Yang, Zhaochun Ren, Yufeng Wang, Chao Chen, Haizhou Sun, Xiaofei Zhu, and Xiangwen Liao. 2024 · 2024
Closest in time.
CTSM: Combining Trait and State Emotions for Empathetic Response Model
Wang Yufeng, Chen Chao, Yang Zhou, Wang Shuhui, and Liao Xiangwen. 2024 · 2024
Closest in time.
Lory: Fully Differentiable Mixture-of-Experts for Autoregressive Language Model Pre-training
Zexuan Zhong, Mengzhou Xia, Danqi Chen, and Mike Lewis. 2024 · 2024
Closest in time.