Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al. 2020 · 2020
Later among the works it cites.
Pegasus: Pre-training with extracted gap-sentences for abstractive summarization
Jingqing Zhang, Yao Zhao, Mohammad Saleh, and Peter Liu. 2020 · 2020
Later among the works it cites.
Assessing political prudence of open-domain chatbots
Yejin Bang, Nayeon Lee, Etsuko Ishii, Andrea Madotto, and Pascale Fung. 2021 · 2021
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Later among the works it cites.
Evaluating large language models trained on code
Original
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al. 2021 · 2021
Later among the works it cites.
Documenting large webtext corpora: A case study on the colossal clean crawled corpus
Jesse Dodge, Maarten Sap, Ana Marasović, William Agnew, Gabriel Ilharco, Dirk Groeneveld, Margaret Mitchell, and Matt Gardner. 2021 · 2021
Later among the works it cites.
Moral stories: Situated reasoning about norms, intents, actions, and their consequences
Denis Emelin, Ronan Le Bras, Jena D Hwang, Maxwell Forbes, and Yejin Choi. 2021 · 2021
Later among the works it cites.
Kgap: Knowledge graph augmented political perspective detection in news media
Original
Shangbin Feng, Zilong Chen, Wenqian Zhang, Qingyao Li, Qinghua Zheng, Xiaojun Chang, and Minnan Luo. 2021 · 2021
Later among the works it cites.
A survey of race, racism, and anti-racism in nlp
Anjalie Field, Su Lin Blodgett, Zeerak Waseem, and Yulia Tsvetkov. 2021 · 2021
Later among the works it cites.
Detecting cross-geographic biases in toxicity modeling on social media
Sayan Ghosh, Dylan Baker, David Jurgens, and Vinodkumar Prabhakaran. 2021 · 2021
Later among the works it cites.
The theory of the political spectrum
Allen Gindler. 2021 · 2021
Later among the works it cites.
Immigration, race & political polarization
Michael Hout and Christopher Maggio. 2021 · 2021
Later among the works it cites.
On transferability of bias mitigation effects in language model fine-tuning
Xisen Jin, Francesco Barbieri, Brendan Kennedy, Aida Mostafazadeh Davani, Leonardo Neves, and Xiang Ren. 2021 · 2021
Later among the works it cites.
The dangers of disinformation
Anti-Defamation League. 2021 · 2021
Later among the works it cites.
Contextualized perturbation for textual adversarial attack
Dianqi Li, Yizhe Zhang, Hao Peng, Liqun Chen, Chris Brockett, Ming-Ting Sun, and Bill Dolan. 2021 · 2021
Later among the works it cites.
Mitigating political bias in language models through reinforced calibration
Ruibo Liu, Chenyan Jia, Jason Wei, Guangxuan Xu, Lili Wang, and Soroush Vosoughi. 2021 · 2021
Later among the works it cites.
Hatexplain: A benchmark dataset for explainable hate speech detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam, Chris Biemann, Pawan Goyal, and Animesh Mukherjee. 2021 · 2021
Later among the works it cites.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Later among the works it cites.
What sounds “right” to me? experiential factors in the perception of political ideology
Qinlan Shen and Carolyn Rose. 2021 · 2021
Later among the works it cites.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
Ben Wang and Aran Komatsuzaki. 2021 · 2021
Later among the works it cites.
Modelling cultural and socio-economic dimensions of political bias in German tweets
Aishwarya Anegundi, Konstantin Schulz, Christian Rauh, and Georg Rehm. 2022 · 2022
Later among the works it cites.
Out of one, many: Using language models to simulate human samples
Original
Lisa P. Argyle, E. Busby, Nancy Fulda, Joshua Ronald Gubler, Christopher Michael Rytting, and David Wingate. 2022 · 2022
Later among the works it cites.
Spinning language models: Risks of propaganda-as-a-service and countermeasures
Eugene Bagdasaryan and Vitaly Shmatikov. 2022 · 2022
Later among the works it cites.
On the intrinsic and extrinsic fairness evaluation metrics for contextualized language representations
Yang Cao, Yada Pruksachatkun, Kai-Wei Chang, Rahul Gupta, Varun Kumar, Jwala Dhamala, and Aram Galstyan. 2022 · 2022
Later among the works it cites.
Dealing with disagreements: Looking beyond the majority vote in subjective annotations
Aida Mostafazadeh Davani, Mark Díaz, and Vinodkumar Prabhakaran. 2022 · 2022
Later among the works it cites.
PAR: Political actor representation learning with social context and expert knowledge
Shangbin Feng, Zhaoxuan Tan, Zilong Chen, Ningnan Wang, Peisheng Yu, Qinghua Zheng, Xiaojun Chang, and Minnan Luo. 2022 · 2022
Later among the works it cites.
Datavoidant: An ai system for addressing political data voids on social media
Claudia Flores-Saviaga, Shangbin Feng, and Saiph Savage. 2022 · 2022
Later among the works it cites.
Exploring the role of grammar and word choice in bias toward african american english (aae) in hate speech classification
Camille Harris, Matan Halevy, Ayanna Howard, Amy Bruckman, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Language generation models can cause harm: So what can we do about it? an actionable survey
Original
Sachin Kumar, Vidhisha Balachandran, Lucille Njoo, Antonios Anastasopoulos, and Yulia Tsvetkov. 2022 · 2022
Later among the works it cites.
Herb: Measuring hierarchical regional bias in pre-trained language models
Yizhi Li, Ge Zhang, Bohao Yang, Chenghua Lin, Anton Ragni, Shi Wang, and Jie Fu. 2022 · 2022
Later among the works it cites.
Gendered mental health stigma in masked language models
Inna Lin, Lucille Njoo, Anjalie Field, Ashish Sharma, Katharina Reinecke, Tim Althoff, and Yulia Tsvetkov. 2022 · 2022
Later among the works it cites.
POLITICS: Pretraining with same-story article comparison for ideology prediction and stance detection
Yujian Liu, Xinliang Frederick Zhang, David Wegsman, Nicholas Beauchamp, and Lu Wang. 2022a · 2022
Later among the works it cites.
POLITICS: Pretraining with same-story article comparison for ideology prediction and stance detection
Yujian Liu, Xinliang Frederick Zhang, David Wegsman, Nicholas Beauchamp, and Lu Wang. 2022b · 2022
Later among the works it cites.
Measuring alignment of online grassroots political communities with political campaigns
Cameron Raymond, Isaac Waller, and Ashton Anderson. 2022 · 2022
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2022 · 2022
Later among the works it cites.
On second thought, let’s not think step by step! bias and toxicity in zero-shot reasoning
Original
Omar Shaikh, Hongxin Zhang, William Held, Michael Bernstein, and Diyi Yang. 2022 · 2022
Later among the works it cites.
Upstream Mitigation Is Not
Ryan Steed, Swetasudha Panda, Ari Kobren, and Michael Wick. 2022 · 2022
Later among the works it cites.
How hate speech varies by target identity: A computational analysis
Michael Yoder, Lynnette Ng, David West Brown, and Kathleen Carley. 2022 · 2022
Later among the works it cites.
KCD: Knowledge walks and textual cues enhanced political perspective detection in news media
Wenqian Zhang, Shangbin Feng, Zilong Chen, Zhenyu Lei, Jundong Li, and Minnan Luo. 2022 · 2022
Later among the works it cites.
Gpt-4 technical report
Original
OpenAI. 2023 · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Original
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.