Fetching the paper…
Reading the bibliography…
Conversational moderation of online communities is crucial to maintaining civility for a constructive environment, but it is challenging to scale and harmful to moderators.
Acute-eval: Improved dialogue evaluation with optimized questions and multi-turn comparisons
Margaret Li, Jason Weston, and Stephen Roller. 2019 · 1909
Earlier work this paper cites.
Behavioral study of obedience
Stanley Milgram. 1963 · 1963
Earlier work this paper cites.
Role playing and interpersonal-conflict reduction
Arthur C Bohart. 1977 · 1977
Earlier work this paper cites.
Interpersonal dynamics in a simulated prison
Craig Haney, Curtis Banks, and Philip Zimbardo. 1973 · 1977
Earlier work this paper cites.
The strategic use of interests, rights, and power to resolve disputes
Anne L Lytle, Jeanne M Brett, and Debra L Shapiro. 1999 · 1999
Earlier work this paper cites.
Demonstrating the power of social situations via a simulated prison experiment
APA. 2004 · 2004
Earlier work this paper cites.
Teaching communication skills using role-play: an experience-based guide for educators
Vicki A Jackson and Anthony L Back. 2011 · 2011
Earlier work this paper cites.
Role of play in social skills and intelligence of children
Aghajani Hashtchin Tahmores. 2011 · 2011
Earlier work this paper cites.
Regulating behavior in online communities
Sara Kiesler, Robert Kraut, Paul Resnick, and Aniket Kittur. 2012 · 2012
Earlier work this paper cites.
The socratic method in cognitive behavioural therapy: A narrative review
Gavin I Clark and Sarah J Egan. 2015 · 2015
Earlier work this paper cites.
The virtues of moderation
James Grimmelmann. 2015 · 2015
Earlier work this paper cites.
Nonviolent communication: A language of life: Life-changing tools for healthy relationships
Marshall B Rosenberg and Deepak Chopra. 2015 · 2015
Earlier work this paper cites.
Quantifying toxicity and verbal violence on twitter
Joshua Guberman, Carol Schmitz, and Libby Hemphill. 2016 · 2016
Earlier work this paper cites.
Effectiveness of a role-play simulation program involving the sbar technique: A quasi-experimental study
Mi Yu and Kyung ja Kang. 2017 · 2017
Earlier work this paper cites.
Identifying and reducing gender bias in word-level language models
Shikha Bordia and Samuel R. Bowman. 2019 · 2019
Earlier work this paper cites.
Crossmod: A cross-community learning-based system to assist reddit moderators
Eshwar Chandrasekharan, Chaitrali Gandhi, Matthew Wortley Mustelier, and Eric Gilbert. 2019 · 2019
Earlier work this paper cites.
Human-machine collaboration for content regulation: The case of reddit automoderator
Shagun Jhaver, Iris Birman, Eric Gilbert, and Amy Bruckman. 2019 · 2019
Earlier work this paper cites.
Polarisation and peacebuilding strategy on digital media platforms: Current strategies and their discontents
Lydia Laurenson. 2019 · 2019
Earlier work this paper cites.
Hate speech detection: Challenges and solutions
Sean MacAvaney, Hao-Ren Yao, Eugene Yang, Katina Russell, Nazli Goharian, and Ophir Frieder. 2019 · 2019
Cited alongside, same era.
Beyond dyadic interactions: Considering chatbots as community members
Joseph Seering, Michal Luria, Geoff Kaufman, and Jessica Hammer. 2019 · 2019
Cited alongside, same era.
Content removal as a moderation strategy: Compliance and other outcomes in the changemyview community
Kumar Bhargav Srinivasan, Cristian Danescu-Niculescu-Mizil, Lillian Lee, and Chenhao Tan. 2019 · 2019
Cited alongside, same era.
Investigating toxicity across multiple reddit communities, users, and moderators
Hind Almerekhi, Supervised by Bernard J Jansen, and co-supervised by Haewoon Kwak. 2020 · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Reconsidering tweets: Intervening during tweet creation decreases offensive content
Matthew Katsaros, Kathy Yang, and Lauren Fratamico. 2022 · 2022
Later among the works it cites.
ProsocialDialog: A prosocial backbone for conversational agents
Hyunwoo Kim, Youngjae Yu, Liwei Jiang, Ximing Lu, Daniel Khashabi, Gunhee Kim, Yejin Choi, and Maarten Sap. 2022 · 2022
Later among the works it cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
Bloom: A 176b-parameter open-access multilingual language model
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Echo chambers on social media: A comparative analysis
Matteo Cinelli, Gianmarco De Francisci Morales, Alessandro Galeazzi, Walter Quattrociocchi, and Michele Starnini. 2020 · 2020
Cited alongside, same era.
Investigating gender bias in language models using causal mediation analysis
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Yaron Singer, and Stuart Shieber. 2020 · 2020
Cited alongside, same era.
Agenda pushing in email to thwart phishing
Hyundong Cho, Genevieve Bartlett, and Marjorie Freedman. 2021 · 2021
Cited alongside, same era.
Towards automatic evaluation of dialog systems: A model-free off-policy evaluation approach
Haoming Jiang, Bo Dai, Mengjiao Yang, Tuo Zhao, and Wei Wei. 2021 · 2021
Cited alongside, same era.
Civil rephrases of toxic texts with self-supervised transformers
Léo Laugier, John Pavlopoulos, Jeffrey Sorensen, and Lucas Dixon. 2021 · 2021
Cited alongside, same era.
Mitigating political bias in language models through reinforced calibration
Ruibo Liu, Chenyan Jia, Jason Wei, Guangxuan Xu, Lili Wang, and Soroush Vosoughi. 2021 · 2021
Cited alongside, same era.
Detecting community sensitive norm violations in online conversations
Chan Young Park, Julia Mendelsohn, Karthik Radhakrishnan, Kinjal Jain, Tushar Kanakagiri, David Jurgens, and Yulia Tsvetkov. 2021 · 2021
Cited alongside, same era.
Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, et al. 2022 · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona Diab, Xian Li, Xi Victoria Lin, et al. 2022 · 2022
Later among the works it cites.
Socratic question generation: A novel dataset, models, and evaluation
Beng Heng Ang, Sujatha Das Gollapalli, and See-Kiong Ng. 2023 · 2023
Closest in time.
Leveraging ai for democratic discourse: Chat interventions can improve online political conversations at scale
Lisa P Argyle, Christopher A Bail, Ethan C Busby, Joshua R Gubler, Thomas Howe, Christopher Rytting, Taylor Sorensen, and David Wingate. 2023 · 2023
Closest in time.
Deep reinforcement learning from human preferences
Paul Christiano, Jan Leike, Tom B. Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2023 · 2023
Closest in time.
SODA: Million-scale dialogue distillation with social commonsense contextualization
Hyunwoo Kim, Jack Hessel, Liwei Jiang, Peter West, Ximing Lu, Youngjae Yu, Pei Zhou, Ronan Bras, Malihe Alikhani, Gunhee Kim, Maarten Sap, and Yejin Choi. 2023 · 2023
Closest in time.
Openassistant conversations–democratizing large language model alignment
Andreas Köpf, Yannic Kilcher, Dimitri von Rütte, Sotiris Anagnostidis, Zhi-Rui Tam, Keith Stevens, Abdullah Barhoum, Nguyen Minh Duc, Oliver Stanley, Richárd Nagyfi, et al. 2023 · 2023
Closest in time.
Analyzing norm violations in live-stream chat
Jihyung Moon, Dong-Ho Lee, Hyundong Cho, Woojeong Jin, Chan Young Park, Minwoo Kim, Jonathan May, Jay Pujara, and Sungjoon Park. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
What if the devil is my guardian angel: Chatgpt as a case study of using chatbots in education
Ahmed Tlili, Boulus Shehata, Michael Agyemang Adarkwah, Aras Bozkurt, Daniel T Hickey, Ronghuai Huang, and Brighter Agyemang. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.
The wisdom of hindsight makes language models better instruction followers
Tianjun Zhang, Fangchen Liu, Justin Wong, Pieter Abbeel, and Joseph E. Gonzalez. 2023 · 2023
Closest in time.
Lima: Less is more for alignment
Chunting Zhou, Pengfei Liu, Puxin Xu, Srini Iyer, Jiao Sun, Yuning Mao, Xuezhe Ma, Avia Efrat, Ping Yu, Lili Yu, et al. 2023 · 2023
Closest in time.