Fetching the paper…
Reading the bibliography…
Users' physical safety is an increasing concern as the market for intelligent systems continues to grow, where unconstrained systems may recommend users dangerous actions that can lead to serious injury.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020a · 1901
Earlier work this paper cites.
A just and comprehensive strategy for using nlp to address online abuse
David Jurgens, Eshwar Chandrasekharan, and Libby Hemphill. 2019 · 1906
Earlier work this paper cites.
Confronting abusive language online: A survey from the ethical and human rights perspective
Svetlana Kiritchenko, Isar Nejadgholi, and Kathleen C. Fraser. 2021 · 2012
Earlier work this paper cites.
Learning from bullying traces in social media
Jun-Ming Xu, Kwang-Sung Jun, Xiaojin Zhu, and Amy Bellmore. 2012 · 2012
Earlier work this paper cites.
Towards a corpus of violence acts in Arabic social media
Ayman Alhelbawy, Poesio Massimo, and Udo Kruschwitz. 2016 · 2016
Earlier work this paper cites.
Desmond Upton Patton, Kathleen McKeown, Owen Rambow, and Jamie Macbeth. 2016 · 2016
Earlier work this paper cites.
The gun violence database: A new task and data set for nlp
Ellie Pavlick, Heng Ji, Xiaoman Pan, and Chris Callison-Burch. 2016 · 2016
Earlier work this paper cites.
The future of free speech, trolls, anonymity and fake news online
Lee Rainie, Janna Quitney Anderson, and Jonathan Albright. 2017 · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
Patient and consumer safety risks when using conversational assistants for medical information: An observational study of siri, alexa, and google assistant
Timothy W Bickmore, Ha Trinh, Stefan Olafsson, Teresa K O’Leary, Reza Asadi, Nathaniel M Rickles, and Ricardo Cruz. 2018 · 2018
Earlier work this paper cites.
Detecting gang-involved escalation on social media using context
Serina Chang, Ruiqi Zhong, Ethan Adams, Fei-Tzin Lee, Siddharth Varia, Desmond Patton, William Frey, Chris Kedzie, and Kathleen McKeown. 2018 · 2018
Earlier work this paper cites.
A knowledge-grounded neural conversation model
Marjan Ghazvininejad, Chris Brockett, Ming-Wei Chang, Bill Dolan, Jianfeng Gao, Wen-tau Yih, and Michel Galley. 2018 · 2018
Earlier work this paper cites.
FEVER: a large-scale dataset for fact extraction and VERification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
TwoWingOS: A two-wing optimization strategy for evidential claim verification
Wenpeng Yin and Dan Roth. 2018 · 2018
Earlier work this paper cites.
Finding microaggressions in the wild: A case for locating elusive phenomena in social media posts
Luke Breitfeller, Emily Ahn, Aldrian Obaja Muis, David Jurgens, and Yulia Tsvetkov. 2019 · 2019
Cited alongside, same era.
Detecting cyberbullying and cyberaggression in social media
Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Emiliano De Cristofaro, Gianluca Stringhini, Athena Vakali, and Nicolas Kourtellis. 2019 · 2019
Cited alongside, same era.
Widening the pipeline in human-guided reinforcement learning with explanation and context-aware data augmentation
Lin Guan, Mudit Verma, Sihang Guo, Ruohan Zhang, and Subbarao Kambhampati. 2020 · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2020 · 2020
Cited alongside, same era.
Enhancing the detection of criminal organizations in mexico using ml and nlp
Javier Osorio and Alejandro Beltran. 2020 · 2020
Cited alongside, same era.
Attributed question answering: Evaluation and modeling for attributed large language models
Bernd Bohnet, Vinh Q. Tran, Pat Verga, Roee Aharoni, Daniel Andor, Livio Baldini Soares, Jacob Eisenstein, Kuzman Ganchev, Jonathan Herzig, Kai Hui, Tom Kwiatkowski, Ji Ma, Jianmo Ni, Tal Schuster, William W. Cohen, Michael Collins, Dipanjan Das, Donald Metzler, Slav Petrov, and Kellie Webster. 2022 · 2022
Closest in time.
Drink bleach or do what now? covid-hera: A study of risk-informed health decision making in the presence of covid-19 misinformation
Arkin Dharawat, Ismini Lourentzou, Alex Morales, and ChengXiang Zhai. 2022 · 2022
Closest in time.
Can language models learn from explanations in context?
Andrew K Lampinen, Ishita Dasgupta, Stephanie CY Chan, Kory Matthewson, Michael Henry Tessler, Antonia Creswell, James L McClelland, Jane X Wang, and Felix Hill. 2022 · 2022
Closest in time.
Safetext: A benchmark for exploring physical safety in language models
Sharon Levy, Emily Allaway, Melanie Subbiah, Lydia Chilton, Desmond Patton, Kathleen McKeown, and William Yang Wang. 2022 · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Approaches to automated detection of cyberbullying: A survey
Semiu Salawu, Yulan He, and Joan A. Lumsden. 2020 · 2020
Cited alongside, same era.
Deep neural network for gender-based violence detection on twitter messages
Carlos M Castorena, Itzel M Abundez, Roberto Alejo, Everardo E Granda-Gutiérrez, Eréndira Rendón, and Octavio Villegas. 2021 · 2021
Cited alongside, same era.
A sentiment analysis and unsupervised learning approach to digital violence against women: Monterrey case
Gregorio Arturo Reyes González and Francisco J Cantu-Ortiz. 2021 · 2021
Cited alongside, same era.
GoldenWind at SemEval-2021 task 5: Orthrus - an ensemble approach to identify toxicity
Marco Palomino, Dawid Grad, and James Bedwell. 2021 · 2021
Cited alongside, same era.
Zero-shot fact verification by claim generation
Liangming Pan, Wenhu Chen, Wenhan Xiong, Min-Yen Kan, and William Yang Wang. 2021 · 2021
Cited alongside, same era.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Cited alongside, same era.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al. 2021 · 2021
Cited alongside, same era.
Jiacheng Liu, Alisa Liu, Ximing Lu, Sean Welleck, Peter West, Ronan Le Bras, Yejin Choi, and Hannaneh Hajishirzi. 2022 · 2022
Closest in time.
Memprompt: Memory-assisted prompt editing with user feedback
Aman Madaan, Niket Tandon, Peter Clark, and Yiming Yang. 2022 · 2022
Closest in time.
Mitigating covertly unsafe text within natural language systems
Alex Mei, Anisha Kabir, Sharon Levy, Melanie Subbiah, Emily Allaway, John Judge, Desmond Patton, Bruce Bimber, Kathleen McKeown, and William Yang Wang. 2022 · 2022
Closest in time.
On the safety of conversational models: Taxonomy, dataset, and benchmark
Hao Sun, Guangxuan Xu, Jiawen Deng, Jiale Cheng, Chujie Zheng, Hao Zhou, Nanyun Peng, Xiaoyan Zhu, and Minlie Huang. 2022 · 2022
Closest in time.
Challenging big-bench tasks and whether chain-of-thought can solve them
Mirac Suzgun, Nathan Scales, Nathanael Schärli, Sebastian Gehrmann, Yi Tay, Hyung Won Chung, Aakanksha Chowdhery, Quoc V Le, Ed H Chi, Denny Zhou, et al. 2022 · 2022
Closest in time.
Rationale-augmented ensembles in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, and Denny Zhou. 2022 · 2022
Closest in time.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Closest in time.
Re3: Generating longer stories with recursive reprompting and revision
Kevin Yang, Yuandong Tian, Nanyun Peng, and Dan Klein. 2022 · 2022
Closest in time.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao. 2022 · 2022
Closest in time.