Fetching the paper…
Reading the bibliography…
Large language models such as ChatGPT exhibit striking political biases.
X-Stance: A Multilingual Multi-Target Dataset for Stance Detection
Jannis Vamvas and Rico Sennrich. 2020 · 2003
Earlier work this paper cites.
Contextual Priming: Where People Vote Affects how they Vote
Jonah Berger, Marc Meredith, and S. Christian Wheeler. 2008 · 2008
Earlier work this paper cites.
Three Case Studies from Switzerland: Smartvote
James Thurman and Urs Gasser. 2009 · 2009
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Defending Against Neural Fake News
Rowan Zellers, Ari Holtzman, Hannah Rashkin, Yonatan Bisk, Ali Farhadi, Franziska Roesner, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
Unsupervised Cross-lingual Representation Learning at Scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Earlier work this paper cites.
Learning to Summarize from Human Feedback
Nisan Stiennon, Long Ouyang, Jeff Wu, Daniel M. Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano. 2020 · 2020
Earlier work this paper cites.
OpinionDigest: A Simple Framework for Opinion Summarization
Yoshihiko Suhara, Xiaolan Wang, Stefanos Angelidis, and Wang-Chiew Tan. 2020 · 2020
Earlier work this paper cites.
TRL: Transformer Reinforcement Learning
Leandro von Werra, Younes Belkada, Lewis Tunstall, Edward Beeching, Tristan Thrush, Nathan Lambert, and Shengyi Huang. 2020 · 2020
Earlier work this paper cites.
Persistent Anti-Muslim Bias in Large Language Models
Abubakar Abid, Maheen Farooqi, and James Zou. 2021 · 2021
Earlier work this paper cites.
Informative and Controllable Opinion Summarization
Reinald Kim Amplayo and Mirella Lapata. 2021 · 2021
Earlier work this paper cites.
Gender and Representation Bias in GPT-3 Generated Stories
Li Lucy and David Bamman. 2021 · 2021
Earlier work this paper cites.
MAUVE: Measuring the Gap Between Neural Text and Human Text using Divergence Frontiers
Krishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun, Sean Welleck, Yejin Choi, and Zaid Harchaoui. 2021 · 2021
Cited alongside, same era.
Changing Personality Traits with the help of a Digital Personality Change Intervention
Mirjam Stieger, Christoph Flückiger, Dominik Rüegger, Tobias Kowatsch, Brent W. Roberts, and Mathias Allemand. 2021 · 2021
Cited alongside, same era.
Fine-tuning Language Models to Find Agreement among Humans with Diverse Preferences
Michiel Bakker, Martin Chadwick, Hannah Sheahan, Michael Tessler, Lucy Campbell-Gillingham, Jan Balaguer, Nat McAleese, Amelia Glaese, John Aslanides, Matt Botvinick, and Christopher Summerfield. 2022 · 2022
Cited alongside, same era.
LORA: Low-Rank Adaptation of Large Language Models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Cited alongside, same era.
CommunityLM: Probing Partisan Worldviews from Language Models
Hang Jiang, Doug Beeferman, Brandon Roy, and Deb Roy. 2022 · 2022
Cited alongside, same era.
Zephyr: Direct Distillation of LM Alignment
Lewis Tunstall, Edward Beeching, Nathan Lambert, Nazneen Rajani, Kashif Rasul, Younes Belkada, Shengyi Huang, Leandro von Werra, Clémentine Fourrier, Nathan Habib, Nathan Sarrazin, Omar Sanseviero, Alexander M. Rush, and Thomas Wolf. 2023 · 2023
Later among the works it cites.
Controlled Text Generation with Natural Language Instructions
Wangchunshu Zhou, Yuchen Eleanor Jiang, Ethan Wilcox, Ryan Cotterell, and Mrinmaya Sachan. 2023 · 2023
Later among the works it cites.
Llama 3 Model Card
AI@Meta. 2024 · 2024
Closest in time.
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
Yejin Bang, Delong Chen, Nayeon Lee, and Pascale Fung. 2024 · 2024
Closest in time.
Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration
Shangbin Feng, Taylor Sorensen, Yuhan Liu, Jillian Fisher, Chan Young Park, Yejin Choi, and Yulia Tsvetkov. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models
Shangbin Feng, Chan Young Park, Yuhan Liu, and Yulia Tsvetkov. 2023 · 2023
Cited alongside, same era.
Jochen Hartmann, Jasper Schwenzow, and Maximilian Witte. 2023 · 2023
Cited alongside, same era.
Co-Writing with Opinionated Language Models Affects Users’ Views
Maurice Jakesch, Advait Bhat, Daniel Buschek, Lior Zalmanson, and Mor Naaman. 2023 · 2023
Cited alongside, same era.
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Stefano Ermon, Christopher D. Manning, and Chelsea Finn. 2023 · 2023
Cited alongside, same era.
The Political Biases of ChatGPT
David Rozado. 2023 · 2023
Cited alongside, same era.
Alpaca: A Strong, Replicable Instruction-following Model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B Hashimoto. 2023 · 2023
Cited alongside, same era.
Closest in time.
ORPO: Monolithic Preference Optimization without Reference Model
Jiwoo Hong, Noah Lee, and James Thorne. 2024 · 2024
Closest in time.
Reinventing Search with a New AI-powered Microsoft Bing and Edge, Your Copilot for the Web
Yusuf Mehdi. 2023 · 2024
Closest in time.
More Human Than Human: Measuring ChatGPT Political Bias
Fabio Motoki, Valdemar Pinho Neto, and Victor Rodrigues. 2024 · 2024
Closest in time.
OpenAI, Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, and Ilge Akkaya et al. 2024 · 2024
Closest in time.
The Self-Perception and Political Biases of ChatGPT
Jérôme Rutinowski, Sven Franke, Jan Endendyk, Ina Dormuth, Moritz Roidl, and Markus Pauly. 2024 · 2024
Closest in time.
Generative Echo Chamber? Effects of LLM-Powered Search Systems on Diverse Information Seeking
Nikhil Sharma, Q. Vera Liao, and Ziang Xiao. 2024 · 2024
Closest in time.