Fetching the paper…
Reading the bibliography…
As large language models (LLMs) become more capable, there is growing excitement about the possibility of using LLMs as proxies for humans in real-world tasks where subjective labels are desired, such as in surveys and opinion polling.
The optimal number of response alternatives for a scale: A review
Eli P Cox III · 1980
Earlier work this paper cites.
Intensity measures of consumer preference
John R Hauser and Steven M Shugan · 1980
Earlier work this paper cites.
Effects of question order on survey responses
Sam G McFarland · 1981
Earlier work this paper cites.
The effect of the question on survey responses: A review
Graham Kalton and Howard Schuman · 1982
Earlier work this paper cites.
Social desirability bias: A demonstration and technique for its reduction
Randall A Gordon · 1987
Earlier work this paper cites.
Response effects in surveys
Hans-J Hippler and Norbert Schwarz · 1987
Earlier work this paper cites.
Response effects in mail surveys
Stephen A Ayidiya and McKee J McClendon · 1990
Earlier work this paper cites.
Acquiescence and recency response-order effects in interview surveys
McKee J McClendon · 1991
Earlier work this paper cites.
A cognitive model of response-order effects in survey measurement
Norbert Schwarz, Hans-J Hippler, and Elisabeth Noelle-Neumann · 1992
Earlier work this paper cites.
An introduction to survey research, polling, and data analysis
Herbert Weisberg, Jon A Krosnick, and Bruce D Bowen · 1996
Earlier work this paper cites.
Questions and answers in attitude surveys: Experiments on question form, wording, and context
Howard Schuman and Stanley Presser · 1996
Earlier work this paper cites.
Do polls reflect opinions or do opinions reflect polls? the impact of political polling on voters’ expectations, preferences, and behavior
Vicki G Morwitz and Carol Pluzinski · 1996
Earlier work this paper cites.
Middle alternatives, acquiescence, and the quality of questionnaire data
Colm A O’Muircheartaigh, Jon A Krosnick, Armin Helic, et al · 2001
Earlier work this paper cites.
Peer reviewed: a catalog of biases in questionnaires
Bernard CK Choi and Anita WP Pak · 2005
Earlier work this paper cites.
The significance of letter position in word recognition
Graham Rawlinson · 2007
Earlier work this paper cites.
Patient satisfaction survey as a tool towards quality improvement
Rashid Al-Abri and Amina Al-Balushi · 2014
Earlier work this paper cites.
Response order effects in the youth tobacco survey: Results of a split-ballot experiment
Alissa O’Halloran, S Sean Hu, Ann Malarcher, Robert McMillen, Nell Valentine, Mary A Moore, Jennifer J Reid, Natalie Darling, and Robert B Gerzoff · 2014
Earlier work this paper cites.
Questionnaire design: How to plan, structure and write survey material for effective market research
Ian Brace · 2018
Earlier work this paper cites.
Universal adversarial triggers for attacking and analyzing NLP
Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh · 2019
Cited alongside, same era.
How can we know what language models know?
Zhengbao Jiang, Frank F Xu, Jun Araki, and Graham Neubig · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Making pre-trained language models better few-shot learners
Tianyu Gao, Adam Fisch, and Danqi Chen · 2021
Cited alongside, same era.
Language models show human-like content effects on reasoning, 2022
Ishita Dasgupta, Andrew K. Lampinen, Stephanie C. Y. Chan, Antonia Creswell, Dharshan Kumaran, James L. McClelland, and Felix Hill · 2022
Cited alongside, same era.
Collateral facilitation in humans and language models
James Michaelov and Benjamin Bergen · 2022
Whose opinions do language models reflect?
Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee, Percy Liang, and Tatsunori Hashimoto · 2023
Closest in time.
Can ai language models replace human participants?
Danica Dillion, Niket Tandon, Yuling Gu, and Kurt Gray · 2023
Closest in time.
Large language models as simulated economic agents: What can we learn from homo silicus?
John J Horton · 2023
Closest in time.
Melanie Sclar, Yejin Choi, Yulia Tsvetkov, and Alane Suhr · 2023
Closest in time.
On large language models’ selection bias in multi-choice questions
Chujie Zheng, Hao Zhou, Fandong Meng, Jie Zhou, and Minlie Huang · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Use-case-grounded simulations for explanation evaluation
Valerie Chen, Nari Johnson, Nicholay Topin, Gregory Plumb, and Ameet Talwalkar · 2022
Cited alongside, same era.
Structural persistence in language models: Priming as a window into abstract language representations
Arabella Sinclair, Jaap Jumelet, Willem Zuidema, and Raquel Fernández · 2022
Cited alongside, same era.
Multitask prompted training enables zero-shot task generalization
Victor Sanh, Albert Webson, Colin Raffel, Stephen Bach, Lintang Sutawika, Zaid Alyafeai, Antoine Chaffin, Arnaud Stiegler, Arun Raja, Manan Dey, M Saiful Bari, Canwen Xu, Urmish Thakker, Shanya Sharma Sharma, Eliza Szczechla, Taewoon Kim, Gunjan Chhablani, Nihal Nayak, Debajyoti Datta, Jonathan Chang, Mike Tian-Jian Jiang, Han Wang, Matteo Manica, Sheng Shen, Zheng Xin Yong, Harshit Pandey, Rachel Bawden, Thomas Wang, Trishala Neeraj, Jos Rozen, Abheesht Sharma, Andrea Santilli, Thibault Fevry, Jason Alan Fries, Ryan Teehan, Teven Le Scao, Stella Biderman, Leo Gao, Thomas Wolf, and Alexander M Rush · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Cited alongside, same era.
Out of One, Many: Using Language Models to Simulate Human Samples
Lisa P. Argyle, Ethan C. Busby, Nancy Fulda, Joshua Gubler, Christopher Rytting, and David Wingate · 2022
Cited alongside, same era.
Ignore previous prompt: Attack techniques for language models
Fábio Perez and Ian Ribeiro · 2022
Cited alongside, same era.
Pouya Pezeshkpour and Estevam Hruschka · 2023
Closest in time.
Inverse scaling: When bigger isn’t better
Ian R. McKenzie, Alexander Lyzhov, Michael Martin Pieler, Alicia Parrish, Aaron Mueller, Ameya Prabhu, Euan McLean, Xudong Shen, Joe Cavanagh, Andrew George Gritsevskiy, Derik Kauffman, Aaron T. Kirtland, Zhengping Zhou, Yuhui Zhang, Sicong Huang, Daniel Wurgaft, Max Weiss, Alexis Ross, Gabriel Recchia, Alisa Liu, Jiacheng Liu, Tom Tseng, Tomasz Korbak, Najoung Kim, Samuel R. Bowman, and Ethan Perez · 2023
Closest in time.
Syntax and semantics meet in the “middle”: Probing the syntax-semantics interface of lms through agentivity
Lindia Tjuatja, Emmy Liu, Lori Levin, and Graham Neubig · 2023
Closest in time.
Towards measuring the representation of subjective global opinions in language models, 2023
Esin Durmus, Karina Nyugen, Thomas I. Liao, Nicholas Schiefer, Amanda Askell, Anton Bakhtin, Carol Chen, Zac Hatfield-Dodds, Danny Hernandez, Nicholas Joseph, Liane Lovitt, Sam McCandlish, Orowa Sikder, Alex Tamkin, Janel Thamkul, Jared Kaplan, Jack Clark, and Deep Ganguli · 2023
Closest in time.
Black box adversarial prompting for foundation models
Natalie Maus, Patrick Chao, Eric Wong, and Jacob R Gardner · 2023
Closest in time.
Universal and transferable adversarial attacks on aligned language models, 2023
Andy Zou, Zifan Wang, J. Zico Kolter, and Matt Fredrikson · 2023
Closest in time.
Are language models worse than humans at following prompts? it’s complicated
Albert Webson, Alyssa Loo, Qinan Yu, and Ellie Pavlick · 2023
Closest in time.
Evaluating Large Language Models in Generating Synthetic HCI Research Data: a Case Study
Perttu Hämäläinen, Mikke Tavast, and Anton Kunnari · 2023
Closest in time.
Chatgpt outperforms crowd workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli · 2023
Closest in time.
Language Models Trained on Media Diets Can Predict Public Opinion, March 2023
Eric Chu, Jacob Andreas, Stephen Ansolabehere, and Deb Roy · 2023
Closest in time.
Ai-augmented surveys: Leveraging large language models for opinion prediction in nationally representative surveys, 2023
Junsol Kim and Byungkyu Lee · 2023
Closest in time.
Evaluating the moral beliefs encoded in LLMs
Nino Scherrer, Claudia Shi, Amir Feder, and David Blei · 2023
Closest in time.
Bridging the Gap: A Survey on Integrating (Human) Feedback for Natural Language Generation
Patrick Fernandes, Aman Madaan, Emmy Liu, António Farinhas, Pedro Henrique Martins, Amanda Bertsch, José G. C. de Souza, Shuyan Zhou, Tongshuang Wu, Graham Neubig, and André F. T. Martins · 2023
Closest in time.