Fetching the paper…
Reading the bibliography…
An increasingly prevalent problem for intelligent technologies is text safety, as uncontrolled systems may generate recommendations to their users that lead to injury or life-threatening consequences.
A just and comprehensive strategy for using nlp to address online abuse
David Jurgens, Eshwar Chandrasekharan, and Libby Hemphill. 2019 · 1906
Earlier work this paper cites.
Universal adversarial triggers for attacking and analyzing nlp
Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh. 2019 · 1908
Earlier work this paper cites.
Fine-tuning language models from human preferences
Daniel M. Ziegler, Nisan Stiennon, Jeff Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving. 2019 · 1909
Earlier work this paper cites.
Plug and play language models: A simple approach to controlled text generation
Sumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung, Eric Frank, Piero Molino, Jason Yosinski, and Rosanne Liu. 2019 · 1912
Earlier work this paper cites.
Upper modeling: A general organization of knowledge for natural language processing
John A. Bateman. 1990 · 1990
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Acquiring correct knowledge for natural language generation
Ehud Reiter, Rohan K. Robertson, and Somayajulu Gowri Sripada. 2003 · 2003
Earlier work this paper cites.
The unified medical language system (umls): integrating biomedical terminology
Olivier Bodenreider. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
On faithfulness and factuality in abstractive summarization
Joshua Maynez, Shashi Narayan, Bernd Bohnet, and Ryan T. McDonald. 2020 · 2005
Earlier work this paper cites.
Neurologic decoding: (un)supervised neural text generation with predicate logic constraints
Ximing Lu, Peter West, Rowan Zellers, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2020 · 2010
Earlier work this paper cites.
Constrained abstractive summarization: Preserving factual consistency with constrained generation
Yuning Mao, Xiang Ren, Heng Ji, and Jiawei Han. 2020 · 2010
Earlier work this paper cites.
Eliciting knowledge from language models using automatically generated prompts
Taylor Shin, Yasaman Razeghi, Robert L Logan IV, Eric Wallace, and Sameer Singh. 2020 · 2010
Earlier work this paper cites.
Counterfactual explanations for machine learning: A review
Sahil Verma, John P. Dickerson, and Keegan E. Hines. 2020 · 2010
Earlier work this paper cites.
Hummod: A modeling environment for the simulation of integrative human physiology
Robert L. Hester, Alison J. Brown, Leland D. Husband, Radu Iliescu, Drew Pruett, Richard L. Summers, and Thomas G. Coleman. 2011 · 2011
Earlier work this paper cites.
Semeval-2012 task 7: Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Andrew S. Gordon, Zornitsa Kozareva, and Melissa Roemmele. 2012 · 2012
Earlier work this paper cites.
Confronting abusive language online: A survey from the ethical and human rights perspective
Svetlana Kiritchenko, Isar Nejadgholi, and Kathleen C. Fraser. 2021 · 2012
Earlier work this paper cites.
Learning from bullying traces in social media
Jun-Ming Xu, Kwang-Sung Jun, Xiaojin Zhu, and Amy Bellmore. 2012 · 2012
Earlier work this paper cites.
Automatic identification of potentially contradictory claims to support systematic reviews
Abdulaziz Alamri and Mark Stevenson. 2015 · 2015
Earlier work this paper cites.
A corpus and cloze evaluation for deeper understanding of commonsense stories
N. Mostafazadeh, Nathanael Chambers, Xiaodong He, Devi Parikh, Dhruv Batra, Lucy Vanderwende, Pushmeet Kohli, and James F. Allen. 2016 · 2016
Earlier work this paper cites.
Desmond Upton Patton, Kathleen McKeown, Owen Rambow, and Jamie Macbeth. 2016 · 2016
Earlier work this paper cites.
The gun violence database: A new task and data set for nlp
Ellie Pavlick, Heng Ji, Xiaoman Pan, and Chris Callison-Burch. 2016 · 2016
Earlier work this paper cites.
Domain aware neural dialog system
Sajal Choudhary, Prerna Srivastava, Lyle H. Ungar, and João Sedoc. 2017 · 2017
Earlier work this paper cites.
European union regulations on algorithmic decision-making and a "right to explanation"
Bryce Goodman and Seth Flaxman. 2017 · 2017
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang. 2017 · 2017
Earlier work this paper cites.
Preclude: Conflict detection in textual health advice
Sarah Masud Preum, Md. Abu Sayeed Mondol, Meiyi Ma, Hongning Wang, and John A. Stankovic. 2017 · 2017
Earlier work this paper cites.
The future of free speech, trolls, anonymity and fake news online
Lee Rainie, Janna Quitney Anderson, and Jonathan Albright. 2017 · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi. 2017 · 2017
Cited alongside, same era.
Understanding abuse: A typology of abusive language detection subtasks
Zeerak Waseem, Thomas Davidson, Dana Warmsley, and Ingmar Weber. 2017 · 2017
Cited alongside, same era.
Flexible end-to-end dialogue system for knowledge grounded conversation
Wenya Zhu, Kaixiang Mo, Yu Zhang, Zhangbin Zhu, Xuezheng Peng, and Qiang Yang. 2017 · 2017
Cited alongside, same era.
Peeking inside the black-box: A survey on explainable artificial intelligence (xai)
Amina Adadi and Mohammed Berrada. 2018 · 2018
Cited alongside, same era.
On the robustness of interpretability methods
David Alvarez-Melis and Tommi S. Jaakkola. 2018 · 2018
Cited alongside, same era.
A sentiment analysis and unsupervised learning approach to digital violence against women: Monterrey case
Gregorio Arturo Reyes González and Francisco J Cantu-Ortiz. 2021 · 2021
Later among the works it cites.
Open-Domain question-Answering for COVID-19 and other emergent domains
Sharon Levy, Kevin Mo, Wenhan Xiong, and William Yang Wang. 2021a · 2021
Later among the works it cites.
Investigating memorization of conspiracy theories in text generation
Sharon Levy, Michael Saxon, and William Yang Wang. 2021b · 2021
Later among the works it cites.
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang. 2021 · 2021
Later among the works it cites.
Neurologic a*esque decoding: Constrained text generation with lookahead heuristics
Ximing Lu, Sean Welleck, Peter West, Liwei Jiang, Jungo Kasai, Daniel Khashabi, Ronan Le Bras, Lianhui Qin, Youngjae Yu, Rowan Zellers, Noah A. Smith, and Yejin Choi. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Serina Chang, Ruiqi Zhong, Ethan Adams, Fei-Tzin Lee, Siddharth Varia, Desmond Patton, William Frey, Chris Kedzie, and Kathleen McKeown. 2018 · 2018
Cited alongside, same era.
Fever: a large-scale dataset for fact extraction and verification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Cited alongside, same era.
Swag: A large-scale adversarial dataset for grounded commonsense inference
Rowan Zellers, Yonatan Bisk, Roy Schwartz, and Yejin Choi. 2018 · 2018
Cited alongside, same era.
Finding microaggressions in the wild: A case for locating elusive phenomena in social media posts
Luke Breitfeller, Emily Ahn, Aldrian Obaja Muis, David Jurgens, and Yulia Tsvetkov. 2019 · 2019
Cited alongside, same era.
Detecting cyberbullying and cyberaggression in social media
Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Emiliano De Cristofaro, Gianluca Stringhini, Athena Vakali, and Nicolas Kourtellis. 2019 · 2019
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Liability for ai decision-making: some legal and ethical considerations
Iria Giuffrida. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Social work in data science: Tech policy gaps and addressing harm
Siva Mathiyazhagan, Shana Kleiner, and Desmond U. Patton. 2021 · 2021
Later among the works it cites.
Annotators with attitudes: How annotator beliefs and identities bias toxic language detection
Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, and Noah A. Smith. 2021 · 2021
Later among the works it cites.
Self-diagnosis and self-debiasing: A proposal for reducing corpus-based bias in nlp
Timo Schick, Sahana Udupa, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
Process for adapting language models to society (palms) with values-targeted datasets
Irene Solaiman and Christy Dennison. 2021 · 2021
Later among the works it cites.
On hallucination and predictive uncertainty in conditional language generation
Yijun Xiao and William Yang Wang. 2021 · 2021
Later among the works it cites.
Yubo Xie and Pearl Pu. 2021 · 2021
Later among the works it cites.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Yuntao Bai, Andy Jones, Kamal Ndousse, Amanda Askell, Anna Chen, Nova DasSarma, Dawn Drain, Stanislav Fort, Deep Ganguli, Tom Henighan, Nicholas Joseph, Saurav Kadavath, Jackson Kernion, Tom Conerly, Sheer El-Showk, Nelson Elhage, Zac Hatfield-Dodds, Danny Hernandez, Tristan Hume, Scott Johnston, Shauna Kravec, Liane Lovitt, Neel Nanda, Catherine Olsson, Dario Amodei, Tom Brown, Jack Clark, Sam McCandlish, Chris Olah, Ben Mann, and Jared Kaplan. 2022 · 2022
Closest in time.
Melatonin and xanax — The Recovery Village
Melissa Carmona. 2022 · 2022
Closest in time.
Milk crate challenge: Why people are taking huge, terrifying falls on social media — CNET
Erin Carson. 2021 · 2022
Closest in time.
Cinnamon challenge dangerous to lungs, new report warns — CBS News
CBS News. 2013 · 2022
Closest in time.
Safetykit: First aid for measuring safety in open-domain conversational systems
Emily Dinan, Gavin Abercrombie, Ari Bergman, Shannon L. Spruit, Dirk Hovy, Y-Lan Boureau, and Verena Rieser. 2022 · 2022
Closest in time.
The link between nicotine and cancer — Verywell Health
Lynne Eldridge. 2021 · 2022
Closest in time.
Bleach and alcohol make chloroform – why you shouldn’t mix disinfectants — Science Notes
Anne Helmenstine. 2020 · 2022
Closest in time.
True: Re-evaluating factual consistency evaluation
Or Honovich, Roee Aharoni, Jonathan Herzig, Hagai Taitelbaum, Doron Kukliansy, Vered Cohen, Thomas Scialom, Idan Szpektor, Avinatan Hassidim, and Y. Matias. 2022 · 2022
Closest in time.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Closest in time.
Safetext: A benchmark for exploring physical safety in language models
Sharon Levy, Emily Allaway, Melanie Subbiah, Lydia Chilton, Desmond Patton, Kathleen McKeown, and William Yang Wang. 2022 · 2022
Closest in time.
Wei Li, Wenhao Wu, Moye Chen, Jiachen Liu, Xinyan Xiao, and Hua Wu. 2022 · 2022
Closest in time.
Black is the new orange: how to determine ai liability
Paulo Henrique Padovan, Clarice Marinho Martins, and Chris Reed. 2022 · 2022
Closest in time.
General first aid for seizures — Epilepsy Foundation
Patty Obsorne Shafer. 2022 · 2022
Closest in time.
A survey of knowledge-enhanced text generation
W. Yu, Wenhao Yu, Chenguang Zhu, Zaitang Li, Zhiting Hu, Qingyun Wang, Heng Ji, and Meng Jiang. 2022 · 2022
Closest in time.
A survey of controllable text generation using transformer-based pre-trained language models
Hanqing Zhang, Haolin Song, Shaoyu Li, Ming Zhou, and Dawei Song. 2022 · 2022
Closest in time.
Users are the north star for ai transparency
Alex Mei, Michael Saxon, Shiyu Chang, Zachary C. Lipton, and William Yang Wang. 2023 · 2023
Closest in time.