Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs),such as ChatGPT, are increasingly used in research, ranging from simple writing assistance to complex data annotation tasks.
Tests of Equality Between Sets of Coefficients in Two Linear Regressions
Gregory C. Chow · 1960
Earlier work this paper cites.
The Problems of LLM-generated Data in Social Science Research
Luca Rossi, Katherine Harrison, and Irina Shklovski · 1971
Earlier work this paper cites.
Prolegomena to any future artificial moral agent
Colin Allen, Gary Varner, and Jason Zinser · 2000
Earlier work this paper cites.
Moral foundations vignettes: a standardized stimulus database of scenarios based on moral foundations theory
Scott Clifford, Vijeth Iyengar, Roberto Cabeza, and Walter Sinnott-Armstrong · 2015
Earlier work this paper cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Earlier work this paper cites.
Off-Duty Deviance in the Eye of the Beholder: Implications of Moral Foundations Theory in the Age of Social Media
Warren Cook and Kristine M. Kuhn · 2021
Earlier work this paper cites.
Do Audiences Judge the Morality of Characters Relativistically? How Interdependence Affects Perceptions of Characters’ Temporal Moral Descent
Matthew Grizzard, Nicholas L Matthews, C Joseph Francemone, and Kaitlin Fitzgerald · 2021
Earlier work this paper cites.
The moral repetition effect: Bad deeds seem less unethical when repeatedly encountered
Daniel A. Effron · 2022
Earlier work this paper cites.
Impression formation stimuli: A corpus of behavior statements rated on morality, competence, informativeness, and believability
Amy Mickelberg, Bradley Walker, Ullrich K. H. Ecker, Piers Howe, Andrew Perfors, and Nicolas Fay · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F Christiano, Jan Leike, and Ryan Lowe · 2022
Earlier work this paper cites.
Potential impact of large language models on academic writing
Fares Alahdab · 2023
Earlier work this paper cites.
Sparks of Artificial General Intelligence: Early experiments with GPT-4, 2023
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, Harsha Nori, Hamid Palangi, Marco Tulio Ribeiro, and Yi Zhang · 2023
Earlier work this paper cites.
Can AI language models replace human participants?
Danica Dillion, Niket Tandon, Yuling Gu, and Kurt Gray · 2023
Earlier work this paper cites.
ChatGPT for good? On opportunities and challenges of large language models for education
Enkelejda Kasneci, Kathrin Sessler, Stefan Küchemann, Maria Bannert, Daryna Dementieva, Frank Fischer, Urs Gasser, Georg Groh, Stephan Günnemann, Eyke Hüllermeier, Stephan Krusche, Gitta Kutyniok, Tilman Michaeli, Claudia Nerdel, Jürgen Pfeffer, Oleksandra Poquet, Michael Sailer, Albrecht Schmidt, Tina Seidel, Matthias Stadler, Jochen Weller, Jochen Kuhn, and Gjergji Kasneci · 2023
Earlier work this paper cites.
Generative Agents: Interactive Simulacra of Human Behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2023
Earlier work this paper cites.
Should artificial intelligent agents be your co-author? Arguments in favour, informed by ChatGPT
Michael Jay Polonsky and Jeffrey D. Rotman · 2023
Earlier work this paper cites.
Language models are not naysayers: An analysis of language models on negation benchmarks, 2023
Thinh Hung Truong, Timothy Baldwin, Karin Verspoor, and Trevor Cohn · 2023
Cited alongside, same era.
Instruction Tuning for Large Language Models: A Survey, 2023
Shengyu Zhang, Linfeng Dong, Xiaoya Li, Sen Zhang, Xiaofei Sun, Shuhe Wang, Jiwei Li, Runyi Hu, Tianwei Zhang, Fei Wu, and Guoyin Wang · 2023
Cited alongside, same era.
Attributions toward artificial agents in a modified Moral Turing Test
Eyal Aharoni, Sharlene Fernandes, Daniel J. Brady, Caelan Alexander, Michael Criner, Kara Queen, Javier Rando, Eddy Nahmias, and Victor Crespo · 2024
Cited alongside, same era.
Exploring the psychology of LLMs’ moral and legal reasoning
Guilherme F.C.F. Almeida, José Luiz Nunes, Neele Engelmann, Alex Wiegmann, and Marcelo De Araújo · 2024
Cited alongside, same era.
Bias and Fairness in Large Language Models: A Survey
Isabel O. Gallegos, Ryan A. Rossi, Joe Barrow, Md Mehrab Tanjim, Sungchul Kim, Franck Dernoncourt, Tong Yu, Ruiyi Zhang, and Nesreen K. Ahmed · 2024
Hallucinations in Large Language Models (LLMs)
G. Pradeep Reddy, Y. V. Pavan Kumar, and K. Purna Prakash · 2024
Later among the works it cites.
Dmitry Scherbakov, Nina Hubig, Vinita Jansari, Alexander Bakumenko, and Leslie A. Lenert · 2024
Later among the works it cites.
Reclaiming AI as a Theoretical Tool for Cognitive Science
Iris Van Rooij, Olivia Guest, Federico Adolfi, Ronald De Haan, Antonina Kolokolova, and Patricia Rich · 2024
Later among the works it cites.
How Reliable is Your Simulator? Analysis on the Limitations of Current LLM-based User Simulators for Conversational Recommendation
Lixi Zhu, Xiaowen Huang, and Jitao Sang · 2024
Later among the works it cites.
ChatGPT in Academic Writing and Publishing: A Comprehensive Guide
Medhat Zohery · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The Moral Turing Test: Evaluating Human-LLM Alignment in Moral Decision-Making, 2024
Basile Garcia, Crystal Qian, and Stefano Palminteri · 2024
Cited alongside, same era.
AI language models cannot replace human research participants
Jacqueline Harding, William D’Alessandro, N. G. Laskowski, and Robert Long · 2024
Cited alongside, same era.
Aligning Generalisation Between Humans and Machines, 2024
Filip Ilievski, Barbara Hammer, Frank van Harmelen, Benjamin Paassen, Sascha Saralajew, Ute Schmid, Michael Biehl, Marianna Bolognesi, Xin Luna Dong, Kiril Gashteovski, Pascal Hitzler, Giuseppe Marra, Pasquale Minervini, Martin Mundt, Axel-Cyrille Ngonga Ngomo, Alessandro Oltramari, Gabriella Pasi, Zeynep G. Saribatur, Luciano Serafini, John Shawe-Taylor, Vered Shwartz, Gabriella Skitalinskaya, Clemens Stachl, Gido M. van de Ven, and Thomas Villmann · 2024
Cited alongside, same era.
Can large language models replace humans in systematic reviews? Evaluating <span style="font-variant:small-caps;">GPT</span> -4’s efficacy in screening and extracting data from peer-reviewed and grey literature in multiple languages
Qusai Khraisha, Sophie Put, Johanna Kappenberg, Azza Warraitch, and Kristin Hadfield · 2024
Cited alongside, same era.
Evaluating large language models in theory of mind tasks
Michal Kosinski · 2024
Cited alongside, same era.
Fine-Grained Detection of Solidarity for Women and Migrants in 155 Years of German Parliamentary Debates
Aida Kostikova, Dominik Beese, Benjamin Paassen, Ole Pütz, Gregor Wiedemann, and Steffen Eger · 2024
Cited alongside, same era.
Fundamental limitations of generative LLMs
Andrei Kucharavy · 2024
Cited alongside, same era.
A foundation model to predict and capture human cognition
Marcel Binz, Elif Akata, Matthias Bethge, Franziska Brändle, Fred Callaway, Julian Coda-Forno, Peter Dayan, Can Demircan, Maria K. Eckstein, Noémi Éltető, Thomas L. Griffiths, Susanne Haridi, Akshay K. Jagadish, Li Ji-An, Alexander Kipnis, Sreejan Kumar, Tobias Ludwig, Marvin Mathony, Marcelo Mattar, Alireza Modirshanechi, Surabhi S. Nath, Joshua C. Peterson, Milena Rmus, Evan M. Russek, Tankred Saanum, Johannes A. Schubert, Luca M. Schulze Buschoff, Nishad Singhi, Xin Sui, Mirko Thalmann, Fabian J. Theis, Vuong Truong, Vishaal Udandarao, Konstantinos Voudouris, Robert Wilson, Kristin Witte, Shuchen Wu, Dirk U. Wulff, Huadong Xiong, and Eric Schulz · 2025
Closest in time.
Centaur: A model without a theory, July 2025
Jeffrey S Bowers, Guillermo Puebla, Sushrut Thorat, Konstantinos Tsetsos, and Casimir Johannes Hendrikus Ludwig · 2025
Closest in time.
AI language model rivals expert ethicist in perceived moral expertise
Danica Dillion, Debanjan Mondal, Niket Tandon, and Kurt Gray · 2025
Closest in time.
Benefits and Risks of Using AI Agents in Research, March 2025
Mohammad Hosseini, Maya Murad, and David Resnik · 2025
Closest in time.
Re-evaluating Theory of Mind evaluation in large language models, 2025
Jennifer Hu, Felix Sosa, and Tomer Ullman · 2025
Closest in time.
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
Lei Huang, Weijiang Yu, Weitao Ma, Weihong Zhong, Zhangyin Feng, Haotian Wang, Qianglong Chen, Weihua Peng, Xiaocheng Feng, Bing Qin, and Ting Liu · 2025
Closest in time.
Investigating machine moral judgement through the Delphi experiment
Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras, Jenny T. Liang, Sydney Levine, Jesse Dodge, Keisuke Sakaguchi, Maxwell Forbes, Jack Hessel, Jon Borchardt, Taylor Sorensen, Saadia Gabriel, Yulia Tsvetkov, Oren Etzioni, Maarten Sap, Regina Rini, and Yejin Choi · 2025
Closest in time.
Interviewing ChatGPT-Generated Personas to Inform Design Decisions
Jemily Rime · 2025
Closest in time.
Combining Psychology with Artificial Intelligence: What could possibly go wrong?
Iris van Rooij and Olivia Guest · 2025
Closest in time.
What Limits LLM-based Human Simulation: LLMs or Our Design?, 2025
Qian Wang, Jiaying Wu, Zhenheng Tang, Bingqiao Luo, Nuo Chen, Wei Chen, and Bingsheng He · 2025
Closest in time.