Fetching the paper…
Reading the bibliography…
Exponentially growing short video platforms (SVPs) face significant challenges in moderating content detrimental to users' mental health, particularly for minors.
How oversight improves member-maintained communities
Dan Cosley, Dan Frankowski, Sara Kiesler, Loren Terveen, and John Riedl · 2005
Earlier work this paper cites.
Gang networks, neighborhoods and holidays: Spatiotemporal patterns in social media
Nibir Bora, Vladimir Zaytsev, Yu-Han Chang, and Rajiv Maheswaran · 2013
Earlier work this paper cites.
Detecting cyberbullying: query terms and techniques
April Kontostathis, Kelly Reynolds, Andy Garron, and Lynne Edwards · 2013
Earlier work this paper cites.
Locate the hate: Detecting tweets against blacks
Irene Kwok and Yuzhou Wang · 2013
Earlier work this paper cites.
Cybercrime detection in online communications: The experimental case of cyberbullying detection in the twitter network
Mohammed Ali Al-Garadi, Kasturi Dewi Varathan, and Sri Devi Ravana · 2016
Earlier work this paper cites.
The rise of social bots
Emilio Ferrara, Onur Varol, Clayton Davis, Filippo Menczer, and Alessandro Flammini · 2016
Earlier work this paper cites.
Analyzing the targets of hate in online social media
Leandro Silva, Mainack Mondal, Denzil Correa, Fabrício Benevenuto, and Ingmar Weber · 2016
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber · 2017
Earlier work this paper cites.
Argument discovery via crowdsourcing
Quoc Viet Hung Nguyen, Chi Thang Duong, Thanh Tam Nguyen, Matthias Weidlich, Karl Aberer, Hongzhi Yin, and Xiaofang Zhou · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand · 2017
Earlier work this paper cites.
Attention is all you need
A Vaswani · 2017
Earlier work this paper cites.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford · 2018
Earlier work this paper cites.
The spread of true and false news online
Soroush Vosoughi, Deb Roy, and Sinan Aral · 2018
Earlier work this paper cites.
Hate speech detection is not as easy as you may think: A closer look at model validation
Aymé Arango, Jorge Pérez, and Barbara Poblete · 2019
Earlier work this paper cites.
Crossmod: A cross-community learning-based system to assist reddit moderators
Eshwar Chandrasekharan, Chaitrali Gandhi, Matthew Wortley Mustelier, and Eric Gilbert · 2019
Earlier work this paper cites.
Hate speech detection on vietnamese social media text using the bidirectional-lstm model
Hang Thi-Thuy Do, Huy Duc Huynh, Kiet Van Nguyen, Ngan Luu-Thuy Nguyen, and Anh Gia-Tuan Nguyen · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin Ming-Wei Chang Kenton and Lee Kristina Toutanova · 2019
Earlier work this paper cites.
Intelligent design of multimedia content in alibaba
Kui-long Liu, Wei Li, Chang-yuan Yang, and Guang Yang · 2019
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Cited alongside, same era.
Volunteer moderators in twitch micro communities: How they get involved, the roles they play, and the emotional labor they experience
Donghee Yvette Wohn · 2019
Cited alongside, same era.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang · 2019
Cited alongside, same era.
A unified taxonomy of harmful content
Michele Banko, Brendon MacKeen, and Laurie Ray · 2020
Cited alongside, same era.
# bigtech@ minors: Social media algorithms personalize minors’ content after a single session, but not for their protection
Martin Hilbert, Drew P Cingel, Jingwen Zhang, Samantha L Vigil, Jane Shawcroft, Haoning Xue, Arti Thakur, and Zubair Shafiq · 2023
Later among the works it cites.
The moderation of contentious content on twitter
Wei Hu · 2023
Later among the works it cites.
Huan Ma, Changqing Zhang, Huazhu Fu, Peilin Zhao, and Bingzhe Wu · 2023
Later among the works it cites.
A holistic approach to undesired content detection in the real world
Todor Markov, Chong Zhang, Sandhini Agarwal, Florentine Eloundou Nekoul, Theodore Lee, Steven Adler, Angela Jiang, and Lilian Weng · 2023
Later among the works it cites.
Content moderation for evolving policies using binary question answering
Sankha Subhra Mullick, Mohan Bhambhani, Suhit Sinha, Akshat Mathur, Somya Gupta, and Jidnya Shah · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
Measuring and characterizing hate speech on news websites
Savvas Zannettou, Mai ElSherief, Elizabeth Belding, Shirin Nilizadeh, and Gianluca Stringhini · 2020
Cited alongside, same era.
Conceptualising contestability: Perspectives on contesting algorithmic decisions
Henrietta Lyons, Eduardo Velloso, and Tim Miller · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
A framework of severity for harmful content online
Morgan Klaus Scheuerman, Jialun Aaron Jiang, Casey Fiesler, and Jed R Brubaker · 2021
Cited alongside, same era.
Exploiting cloze-questions for few-shot text classification and natural language inference
Timo Schick and Hinrich Schütze · 2021
Cited alongside, same era.
A comparative study of using pre-trained language models for toxic comment classification
Zhixue Zhao, Ziqi Zhang, and Frank Hopfgartner · 2021
Cited alongside, same era.
Later among the works it cites.
Detecting rumours with latency guarantees using massive streaming data
Thanh Tam Nguyen, Thanh Trung Huynh, Hongzhi Yin, Matthias Weidlich, Thanh Thi Nguyen, Thai Son Mai, and Quoc Viet Hung Nguyen · 2023
Later among the works it cites.
Ethical scaling for content moderation: Extreme speech and the (in) significance of artificial intelligence
Sahana Udupa, Antonis Maronikolakis, and Axel Wisiorek · 2023
Later among the works it cites.
Multi-modal misinformation detection: Approaches, challenges and opportunities
Sara Abdali, Sina Shaham, and Bhaskar Krishnamachari · 2024
Later among the works it cites.
Legal and ethical challenges of content moderation: Balancing privacy and free speech in the ai era
Palwasha Bibi · 2024
Later among the works it cites.
Evlm: An efficient vision-language model for visual understanding
Kaibing Chen, Dong Shen, Hanwen Zhong, Huasong Zhong, Kui Xia, Di Xu, Wei Yuan, Yifei Hu, Bin Wen, Tianke Zhang, et al · 2024
Later among the works it cites.
Zhe Chen, Weiyun Wang, Yue Cao, Yangzhou Liu, Zhangwei Gao, Erfei Cui, Jinguo Zhu, Shenglong Ye, Hao Tian, Zhaoyang Liu, et al · 2024
Later among the works it cites.
You only prompt once: On the capabilities of prompt learning on large language models to tackle toxic content
Xinlei He, Savvas Zannettou, Yun Shen, and Yang Zhang · 2024
Later among the works it cites.
Aaron Hurst, Adam Lerer, Adam P Goucher, Adam Perelman, Aditya Ramesh, Aidan Clark, AJ Ostrow, Akila Welihinda, Alan Hayes, Alec Radford, et al · 2024
Later among the works it cites.
Harmful youtube video detection: A taxonomy of online harm and mllms as alternative annotators
Claire Wonjeong Jo, Miki Wesołowska, and Magdalena Wojcieszak · 2024
Later among the works it cites.
Artificial intelligence in e-commerce: a comprehensive analysis
Meet Ashokkumar Joshi · 2024
Later among the works it cites.
Watch your language: Investigating content moderation with large language models
Deepak Kumar, Yousef Anees AbuHashem, and Zakir Durumeric · 2024
Later among the works it cites.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al · 2025
Closest in time.
Civil law (legal system) — Wikipedia, The Free Encyclopedia, 2025
Wikipedia contributors · 2025
Closest in time.
Common law — Wikipedia, The Free Encyclopedia, 2025
Wikipedia contributors · 2025
Closest in time.