Fetching the paper…
Reading the bibliography…
The advanced role-playing capabilities of Large Language Models (LLMs) have enabled rich interactive scenarios, yet existing research in social interactions neglects hallucination while struggling with poor generalizability and implicit character fidelity judgments.
The public acquisition of commonsense knowledge
Push Singh et al. 2002 · 2002
Earlier work this paper cites.
Social exchange theory: An interdisciplinary review
Russell Cropanzano and Marie S Mitchell. 2005 · 2005
Earlier work this paper cites.
Dbpedia: A nucleus for a web of open data
Sören Auer, Christian Bizer, Georgi Kobilarov, Jens Lehmann, Richard Cyganiak, and Zachary Ives. 2007 · 2007
Earlier work this paper cites.
Impression management theory and social psychological research
James T Tedeschi. 2013 · 2013
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi. 2017 · 2017
Earlier work this paper cites.
Degree centrality, betweenness centrality, and closeness centrality in social network
Junlong Zhang and Yu Luo. 2017 · 2017
Earlier work this paper cites.
Geoffrey Irving, Paul Christiano, and Dario Amodei. 2018 · 2018
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al. 2020 · 2020
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Earlier work this paper cites.
Retrieval augmentation reduces hallucination in conversation
Kurt Shuster, Spencer Poff, Moya Chen, Douwe Kiela, and Jason Weston. 2021 · 2021
Earlier work this paper cites.
Progressive generation of long text with pretrained language models
Bowen Tan, Zichao Yang, Maruan Al-Shedivat, Eric Xing, and Zhiting Hu. 2021 · 2021
Earlier work this paper cites.
Measuring progress on scalable oversight for large language models
Samuel R Bowman, Jeeyoon Hyun, Ethan Perez, Edwin Chen, Craig Pettit, Scott Heiner, Kamilė Lukošiūtė, Amanda Askell, Andy Jones, Anna Chen, et al. 2022 · 2022
Earlier work this paper cites.
Extracting latent steering vectors from pretrained language models
Nishant Subramani, Nivedita Suresh, and Matthew E Peters. 2022 · 2022
Earlier work this paper cites.
Chatmatch: Evaluating chatbots by autonomous chat tournaments
Ruolan Yang, Zitong Li, Haifeng Tang, and Kenny Zhu. 2022 · 2022
Earlier work this paper cites.
How would stance detection techniques evolve after the launch of chatgpt?
Bowen Zhang, Daijun Ding, Liwen Jing, Genan Dai, and Nan Yin. 2022 · 2022
Earlier work this paper cites.
Large language models meet harry potter: A dataset for aligning dialogue agents with characters
Nuo Chen, Yan Wang, Haiyun Jiang, Deng Cai, Yuhan Li, Ziyang Chen, Longyue Wang, and Jia Li. 2023 · 2023
Earlier work this paper cites.
Pippa: A partially synthetic conversational dataset
Tear Gosling, Alpin Dale, and Yinhe Zheng. 2023 · 2023
Earlier work this paper cites.
Lei Huang, Weijiang Yu, Weitao Ma, Weihong Zhong, Zhangyin Feng, Haotian Wang, Qianglong Chen, Weihua Peng, Xiaocheng Feng, Bing Qin, et al. 2023 · 2023
Earlier work this paper cites.
Moelora: An moe-based parameter efficient fine-tuning method for multi-task medical applications
Qidong Liu, Xian Wu, Xiangyu Zhao, Yuanshao Zhu, Derong Xu, Feng Tian, and Yefeng Zheng. 2023 · 2023
Earlier work this paper cites.
Introducing chatgpt
OpenAI. 2023 · 2023
Cited alongside, same era.
Character-llm: A trainable agent for role-playing
Yunfan Shao, Linyang Li, Junqi Dai, and Xipeng Qiu. 2023 · 2023
Cited alongside, same era.
Towards understanding sycophancy in language models
Mrinank Sharma, Meg Tong, Tomasz Korbak, David Duvenaud, Amanda Askell, Samuel R Bowman, Esin DURMUS, Zac Hatfield-Dodds, Scott R Johnston, Shauna M Kravec, et al. 2023 · 2023
Cited alongside, same era.
Building persona consistent dialogue agents with offline reinforcement learning
Ryan Shea and Zhou Yu. 2023 · 2023
Cited alongside, same era.
Roleeval: A bilingual role evaluation benchmark for large language models
Tianhao Shen, Sun Li, Quan Tu, and Deyi Xiong. 2023 · 2023
Cited alongside, same era.
Describe, explain, plan and select: interactive planning with large language models enables open-world multi-task agents
Large language models are superpositions of all characters: Attaining arbitrary role-play via self-alignment
Keming Lu, Bowen Yu, Chang Zhou, and Jingren Zhou. 2024 · 2024
Closest in time.
Fine-tuning or retrieval? comparing knowledge injection in llms
Oded Ovadia, Menachem Brief, Moshik Mishaeli, and Oren Elisha. 2024 · 2024
Closest in time.
I learn better if you speak my language: Enhancing large language model fine-tuning with style-aligned response adjustments
Xuan Ren, Biao Wu, and Lingqiao Liu. 2024 · 2024
Closest in time.
Steering llama 2 via contrastive activation addition
Nina Rimsky, Nick Gabrieli, Julian Schulz, Meg Tong, Evan Hubinger, and Alexander Turner. 2024 · 2024
Closest in time.
Mitigating hallucination in fictional character role-play
Nafis Sadeq, Zhouhang Xie, Byungkyu Kang, Prarit Lamba, Xiang Gao, and Julian McAuley. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zihao Wang, Shaofei Cai, Guanzhou Chen, Anji Liu, Xiaojian Ma, Yitao Liang, and Team CraftJarvis. 2023 · 2023
Cited alongside, same era.
Simple synthetic data reduces sycophancy in large language models
Jerry Wei, Da Huang, Yifeng Lu, Denny Zhou, and Quoc V Le. 2023 · 2023
Cited alongside, same era.
Investigating chain-of-thought with chatgpt for stance detection on social media
Bowen Zhang, Xianghua Fu, Daijun Ding, Hu Huang, Yangyang Li, and Liwen Jing. 2023 · 2023
Cited alongside, same era.
Prompt leakage effect and defense strategies for multi-turn llm interactions
Divyansh Agarwal, Alexander R. Fabbri, Ben Risher, Philippe Laban, Shafiq Joty, and Chien-Sheng Wu. 2024 · 2024
Cited alongside, same era.
Timechara: Evaluating point-in-time character hallucination of role-playing large language models
Jaewoo Ahn, Taehyun Lee, Junyoung Lim, Jin-Hwa Kim, Sangdoo Yun, Hwaran Lee, and Gunhee Kim. 2024 · 2024
Cited alongside, same era.
Socialbench: Sociality evaluation of role-playing conversational agents
Hongzhan Chen, Hehong Chen, Ming Yan, Wenshen Xu, Gao Xing, Weizhou Shen, Xiaojun Quan, Chenliang Li, Ji Zhang, and Fei Huang. 2024a · 2024
Cited alongside, same era.
A survey on in-context learning
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Jingyuan Ma, Rui Li, Heming Xia, Jingjing Xu, Zhiyong Wu, Baobao Chang, Xu Sun, Lei Li, and Zhifang Sui. 2024 · 2024
Cited alongside, same era.
Alexander Spangher, Nanyun Peng, Sebastian Gehrmann, and Mark Dredze. 2024 · 2024
Closest in time.
Analyzing the generalization and reliability of steering vectors
Daniel Chee Hian Tan, David Chanin, Aengus Lynch, Adrià Garriga-Alonso, Dimitrios Kanoulas, Brooks Paige, and Robert Kirk · 2024
Closest in time.
Charactereval: A chinese benchmark for role-playing conversational agent evaluation
Quan Tu, Shilong Fan, Zihang Tian, and Rui Yan. 2024 · 2024
Closest in time.
Self-preference bias in llm-as-a-judge
Koki Wataoka, Tsubasa Takahashi, and Ryokan Ri. 2024 · 2024
Closest in time.
From role-play to drama-interaction: An llm solution
Weiqi Wu, Hongqiu Wu, Lai Jiang, Xingyuan Liu, Jiale Hong, Hai Zhao, and Min Zhang. 2024 · 2024
Closest in time.
An Yang, Baosong Yang, Binyuan Hui, Bo Zheng, Bowen Yu, Chang Zhou, Chengpeng Li, Chengyuan Li, Dayiheng Liu, Fei Huang, et al. 2024 · 2024
Closest in time.
Neeko: Leveraging dynamic lora for efficient multi-character role-playing agent
Xiaoyan Yu, Tongxu Luo, Yifan Wei, Fangyu Lei, Yiming Huang, Peng Hao, and Liehuang Zhu. 2024 · 2024
Closest in time.
Characterglm: Customizing chinese conversational ai characters with large language models
Jinfeng Zhou, Zhuang Chen, Dazhen Wan, Bosi Wen, Yi Song, Jifan Yu, Yongkang Huang, Libiao Peng, Jiaming Yang, Xiyao Xiao, et al. 2023b · 2024
Closest in time.
Investigating generalization of one-shot llm steering vectors
Jacob Dunefsky and Arman Cohan. 2025 · 2025
Closest in time.
Shakespearean sparks: The dance of hallucination and creativity in llms’ decoding layers
Zicong He, Boxuan Zhang, and Lu Cheng. 2025 · 2025
Closest in time.
Cross-model transferability among large language models on the platonic representations of concepts
Youcheng Huang, Chen Huang, Duanyu Feng, Wenqiang Lei, and Jiancheng Lv. 2025 · 2025
Closest in time.
Diffsensei: Bridging multi-modal llms and diffusion models for customized manga generation
Jianzong Wu, Chao Tang, Jingbo Wang, Yanhong Zeng, Xiangtai Li, and Yunhai Tong. 2025 · 2025
Closest in time.
Hallucinations can improve large language models in drug discovery
Shuzhou Yuan and Michael Färber. 2025 · 2025
Closest in time.