Fetching the paper…
Reading the bibliography…
One trending application of LLM (large language model) is to use it for content moderation in online platforms.
“Positivism and the Separation of Law and Morals”
H… Hart · 1958
Earlier work this paper cites.
“Hard Cases”
Ronald Dworkin · 1975
Earlier work this paper cites.
“Legal Theory and the Obligation of a Judge: The Hart/Dworkin Dispute”
E. Soper · 1977
Earlier work this paper cites.
“Marshall v. Jerrico, Inc., 446 U.S. 238”, 1980
Thurgood Marshall · 1980
Earlier work this paper cites.
“A Hard Look at Hard Cases: The Nightmare of a Noble Dreamer”
Alan Hutchinson and John Wakefield · 1982
Earlier work this paper cites.
“Easy Cases”
Frederick Schauer · 1985
Earlier work this paper cites.
“Democracy and Determinacy: An Essay on Legal Interpretation”
Allan. Hutchinson · 1989
Earlier work this paper cites.
“Policy Legitimacy and the Supreme Court: The Sources and Contexts of Legitimation”
Jeffery Mondak · 1994
Earlier work this paper cites.
“Law and Disagreement”
Jeremy Waldron · 1999
Earlier work this paper cites.
“The Private Role in Public Governance”
Jody Freeman · 2000
Earlier work this paper cites.
“Procedural Justice”
Lawrence. Solum · 2004
Earlier work this paper cites.
“Handbook of Inter-Rater Reliability: The Definitive Guide to Measuring the Extent of Agreement Among Raters”
Kilem Gwet · 2014
Earlier work this paper cites.
“The Impact of Psychological Science on Policing in the United States: Procedural Justice, Legitimacy, and Effective Law Enforcement”
Tom Tyler, Phillip Goff and Robert. MacCoun · 2015
Earlier work this paper cites.
“Facebook backs down from ’napalm girl’ censorship and reinstates photo”
Sam et. al · 2016
Earlier work this paper cites.
“Fury over Facebook ’Napalm girl’ censorship”, 2016
Zoe Kleinman · 2016
Earlier work this paper cites.
“’Plausible Cause’: Explanatory Standards in the Age of Powerful Machines”
Kiel Brennan-Marquez · 2017
Earlier work this paper cites.
“Custodians of the Internet: Platforms, Content Moderation, and the Hidden Decisions That Shape Social Media”
Tarleton Gillespie · 2018
Earlier work this paper cites.
“The New Governors: The People, Rules, and Processes Governing Online Speech”
Kate Klonick · 2018
Earlier work this paper cites.
“Artificial Intelligence & Human Rights: Opportunities & Risks”, 2018
Filippo Raso et al · 2018
Earlier work this paper cites.
“Digital Constitutionalism: Using the Rule of Law to Evaluate the Legitimacy of Governance by Platforms”
Nicolas Suzor · 2018
Earlier work this paper cites.
“Evaluating the legitimacy of platform governance: A review of research and a shared research agenda”
Nicolas Suzor, Tess Geelen and Sarah West · 2018
Earlier work this paper cites.
“Report Of The Facebook Data Transparency Advisory Group”, 2019
Ben et. al · 2019
Earlier work this paper cites.
“’Did You Suspect the Post Would be Removed?’: Understanding User Reactions to Content Removals on Reddit”
Shagun et. al · 2019
Earlier work this paper cites.
“Probably Speech, Maybe Free: Toward a Probabilistic Understanding of Online Expression and Platform Governance”, 2019
Mike Ananny · 2019
Earlier work this paper cites.
“The Procedure Fetish”
Nicholas Bagley · 2019
Earlier work this paper cites.
“Global Platform Governance: Private Power in the Shadow of the State”
Hannah Bloch-Wehba · 2019
Earlier work this paper cites.
“Not-So-Easy Cases”
Jonathan Crowe · 2019
Earlier work this paper cites.
“Does Transparency in Moderation Really Matter?: User Behavior After Content Removal Explanations on Reddit”
Shagun Jhaver, Amy Bruckman and Eric Gilbert · 2019
Earlier work this paper cites.
“Binary Governance: Lessons from the GDPR’s Approach to Algorithmic Accountability”
Margot. Kaminski · 2019
Earlier work this paper cites.
“Masnick’s Impossibility Theorem: Content Moderation At Scale Is Impossible To Do Well”
Mike Masnick · 2019
Earlier work this paper cites.
“What Do We Mean When We Talk About Transparency? Toward Meaningful Transparency in Commercial Content Moderation”
Nicolas Suzor, Sarah West, Andrew Quodling and Jillian York · 2019
Earlier work this paper cites.
“OVERSIGHT BOARD OVERTURNS FACEBOOK DECISION IN NAZI QUOTE CASE”, 2020
Facebook Board · 2020
Earlier work this paper cites.
“Donation of Pharmaceutical Drugs to Sri Lanka”, 2020
Meta Board · 2020
Earlier work this paper cites.
“The Brussels Effect: How the European Union Rules the World”
Anu Bradford · 2020
Earlier work this paper cites.
“Mark Zuckerberg’s reversal on Holocaustdenial is a 180”, 2020
Elizabeth Dwoskin · 2020
Earlier work this paper cites.
“Content moderation, AI, and the question of scale”
Tarleton Gillespie · 2020
Earlier work this paper cites.
“Mediating Community AI Interaction through Situated Explanation: The Case of AI Led Moderation”
Yubo Kou and Xinning Gui · 2020
Earlier work this paper cites.
“Artificial Intelligence, Content Moderation, and Freedom of Expression”, 2020
Emma Llansó, Joris van Hoboken, Paddy Leerssen and Jaron Harambam · 2020
Earlier work this paper cites.
“How Facebook uses super-efficient AI models to detect hate speech”, 2020
Meta · 2020
Earlier work this paper cites.
“Intercoder Reliability in Qualitative Research: Debates and Practical Guidelines”
Cliodhna O’Connor and Helene Joffe · 2020
Earlier work this paper cites.
“Re-humanizing the platform: Content moderators and the logic of care”
Minna Ruckenstein and Linda Turunen · 2020
Earlier work this paper cites.
“The impact of algorithms for online content filtering or moderation”, 2020
Giovanni Sartor and Andrea Loreggia · 2020
Earlier work this paper cites.
“The Santa Clara Principles”, 2021
Access et. al · 2021
Earlier work this paper cites.
“HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection”
Binny et. al · 2021
Earlier work this paper cites.
“Contestability For Content Moderation”
Kristen et. al · 2021
Earlier work this paper cites.
“Ethical and social risks of harm from Language Models”, 2021
Laura et. al · 2021
Earlier work this paper cites.
“PRO-NAVALNY PROTESTS IN RUSSIA”, 2021
Facebook Board · 2021
Earlier work this paper cites.
Remi Denton et al · 2021
Earlier work this paper cites.
“Governing Online Speech: From ’Posts-As-Trumps’ to Proportionality and Probability”
Evelyn Douek · 2021
Earlier work this paper cites.
“How Many Cases Are Easy?”
Joshua. Fischman · 2021
Earlier work this paper cites.
“The Oversight of Content Moderation by AI: Impact Assessments and Their Limitations”
Yifat Nahmias and Maayan Perel · 2021
Earlier work this paper cites.
“Content Moderation as Systems Thinking”
Evelyn Douek · 2022
Earlier work this paper cites.
“Detection and moderation of detrimental content on social media platforms: current status and future directions”
Vaishali. Gongane, Mousami. Munot1 and Alwin. Anuse · 2022
Earlier work this paper cites.
“Content-filtering AI systems–limitations, challenges and regulatory approaches”
Althaf Marsoof, Andrés Luco, Harry Tan and Shafiq Joty · 2022
Earlier work this paper cites.
“How Meta prioritizes content for review”, 2022
Meta · 2022
Earlier work this paper cites.
“How review teams work”, 2022
Meta · 2022
Earlier work this paper cites.
“When AI moderates online content: effects of human collaboration and interactive transparency on user trust”
Maria. Molina and S. Sundar · 2022
Cited alongside, same era.
“Digital Services Act (Directive 2000/31/EC)”, 2022
European Union · 2022
Cited alongside, same era.
“Attention Is All You Need”, 2023
Ashish et. al · 2023
Cited alongside, same era.
“IFAN: An Explainability-Focused Interaction Framework for Humans and NLP Models”, 2023
Edoardo et. al · 2023
Cited alongside, same era.
“Evaluating GPT-3 generated explanations for hateful content moderation”
Han et. al · 2023
Cited alongside, same era.
“Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models”, 2024
Nishant et. al · 2024
Closest in time.
Nouar et. al · 2024
Closest in time.
Ola et. al · 2024
Closest in time.
“Demystifying ChatGPT: An In -
Pronaya et. al · 2024
Closest in time.
“Large Language Models are Geographically Biased”
Rohin et. al · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Laura et. al · 2023
Cited alongside, same era.
Miao et. al · 2023
Cited alongside, same era.
“Towards Understanding Sycophancy in Language Models”, 2023
Mrinank et. al · 2023
Cited alongside, same era.
“Can Large Language Models Capture Dissenting Human Voices?”
Noah et. al · 2023
Cited alongside, same era.
“Content Moderation for Evolving Policies using Binary Question Answering”
Sankha et. al · 2023
Cited alongside, same era.
“Siren’s Song in the AI Ocean: A Survey on Hallucination in Large Language Models”, 2023
Yue et. al · 2023
Cited alongside, same era.
“What is AI anyway, and why should I care?”, 2023
Marissa D’Alonzo · 2023
Cited alongside, same era.
Rongxin et. al · 2024
Closest in time.
“MisinfoEval: Generative AI in the Era of ’Alternative Facts”’
Saadia et. al · 2024
Closest in time.
“Hate Personified: Investigating the role of LLMs in content moderation”
Sarah et. al · 2024
Closest in time.
“The Explanation That Hits Home: The Characteristics of Verbal Explanations That Affect Human Perception in Subjective Decision-Making”
Sharon et. al · 2024
Closest in time.
“Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers”
Siddique et. al · 2024
Closest in time.
“LLM-based Rewriting of Inappropriate Argumentation using Reinforcement Learning from Machine Feedback”
Timon et. al · 2024
Closest in time.
“Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators through a User-Centric Method”
Yang et. al · 2024
Closest in time.
Yaqiong et. al · 2024
Closest in time.
“Enhancing LLM-based Hatred and Toxicity Detection with Meta-Toxic Knowledge Graph”, 2024
Yibo et. al · 2024
Closest in time.
Yiwen et. al · 2024
Closest in time.
“Can LLM Graph Reasoning Generalize beyond Pattern Memorization?”
Yizhuo et. al · 2024
Closest in time.
“ImpScore: A Learnable Metric For Quantifying The Implicitness Level of Language”, 2024
Yuxin et. al · 2024
Closest in time.
“MLLM-as-a-Judge for Image Safety without Human Labeling”, 2024
ZhentingWang et. al · 2024
Closest in time.
“Hallucination is Inevitable: An Innate Limitation of Large Language Models”, 2024
Ziwei et. al · 2024
Closest in time.
“Transformer – based models for combating rumours on microblogging platforms: a review”
Rini Anggrainingsih, Ghulam Hassan and Amitava Datta · 2024
Closest in time.
“Is Generative AI the Answer for theFailures of Content Moderation?”, 2024
Paul. Barrett and Justin Hendrix · 2024
Closest in time.
“Cycles of Thought: Measuring LLM Confidence through Stable Explanations”, 2024
Evan Becker and Stefano Soatto · 2024
Closest in time.
“Using LLMs to Moderate Content: Are TheyReady for Commercial Use?”, 2024
Alyssa Boicel · 2024
Closest in time.
“Quantifying Uncertainty in Answers from any Language Model and Enhancing their Trustworthiness”
Jiuhai Chen and Jonas Mueller · 2024
Closest in time.
“Automated Claim Matching with Large Language Models: Empowering Fact-Checkers in the Fight Against Misinformation”
Eun Choi and Emilio Ferrara · 2024
Closest in time.
“LLMs for Cyber Security: New Opportunities”, 2024
Dinil Divakaran and Sai Peddinti · 2024
Closest in time.
Bastián González-Bustamante · 2024
Closest in time.
“Classification: Accuracy, recall, precision, and related metrics”, 2024
Google · 2024
Closest in time.
“Stance Detection on Social Media with Fine-Tuned Large Language Models”, 2024
Ilker Gül, Rémi Lebret and Karl Aberer · 2024
Closest in time.
Sahas Koka, Anthony Vuong and Anish Kataria · 2024
Closest in time.
“LLM-Mod: Can Large Language Models Assist Content Moderation?”
Mahi Kolla, Siddharth Salunkhe, Eshwar Chandrasekharan and Koustuv Saha · 2024
Closest in time.
“Watch Your Language: Investigating Content Moderation with Large Language Models”, 2024
Deepak Kumar, Yousef AbuHashem and Zakir Durumeric · 2024
Closest in time.
“Properties and Challenges of LLM-Generated Explanations”
Jenny Kunz and Marco Kuhlmann · 2024
Closest in time.
““HOT” ChatGPT:The Promise of ChatGPT in Detecting and Discriminating Hateful, Offensive, and Toxic Comments on Social Media”
Lingyao Li, Lizhou Fan, Shubham Atreja and Libby Hemphill · 2024
Closest in time.
“Robustifying Safety-Aligned Large Language Models through Clean Data Curation”, 2024
Xiaoqun Liu, Jiacheng Liang, Muchao Ye and Zhaohan Xi · 2024
Closest in time.
Huan Ma et al · 2024
Closest in time.
Huy Nghiem and Halé III · 2024
Closest in time.
“Interpretable Hate Speech Detection via Large Language Model-extracted Rationales”, 2024
Ayushi Nirmal · 2024
Closest in time.
“ChatGPT, can you solve the content moderation dilemma?”
Emmanuel Penagos · 2024
Closest in time.
“Towards Efficient and Explainable Hate Speech Detection via Model Distillation”, 2024
Paloma Piot and Javier Parapar · 2024
Closest in time.
“The perils and promises of fact-checking with large language models”
Dorian Quelle and Alexandre Bovet · 2024
Closest in time.
“Hate speech detection in social media: Techniques, recent trends, and future challenges”
Anchal Rawat, Santosh Kumar and Surender Samant · 2024
Closest in time.
“MBIAS: Mitigating Bias in Large Language Models While Retaining Context”, 2024
Shaina Raza, Ananya Raval and Veronica Chatrath · 2024
Closest in time.
“Multimodal Hate Speech Detection using Fine-Tuned Llama 2 Model”
Keerthana Sasidaran and Geetha J · 2024
Closest in time.
“Content Moderation using AI”, 2024
Swarnim Sawane · 2024
Closest in time.
“Causality Guided Disentanglement for Cross-Platform Hate Speech Detection”
Paras Sheth et al · 2024
Closest in time.
“ChatGPT, AI Large Language Models, And Law”
Harry Surden · 2024
Closest in time.
“Human-LLM Collaborative Annotation Through Effective Verification of LLM Labels”
Xinru Wang et al · 2024
Closest in time.
“Using LLMs for Policy-Driven Content Classification”, 2024
Dave Willner and Samidh Chakrabarti · 2024
Closest in time.
“LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples”, 2024
Jia-Yu Yao et al · 2024
Closest in time.
“From Detection to Explanation: Effective Learning Strategies for LLMs in Online Abusive Language Research”
Chiara et. al · 2025
Closest in time.
“Goodbye human annotators? Content analysis of social policy debates using ChatGPT”
Erwin et. al · 2025
Closest in time.
“U-GIFT: Uncertainty-Guided Firewall for Toxic Speech in Few-Shot Scenario”, 2025
Jiaxin et. al · 2025
Closest in time.
“A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions”
Lei et. al · 2025
Closest in time.
“Silver Lining in the Fake News Cloud: Can Large Language Models Help Detect Misinformation?”
Raghvendra et. al · 2025
Closest in time.