Fetching the paper…
Reading the bibliography…
We propose a novel framework for cross-lingual content flagging with limited target-language data, which significantly outperforms prior work in terms of predictive performance.
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
MBT: A Memory-Based Part of Speech Tagger-Generator
Walter Daelemans, Jakub Zavrel, Peter Berck, and Steven Gillis. 1996 · 1996
Earlier work this paper cites.
Lukas Stappen, Fabian Brunn, and Björn Schuller. 2020 · 2004
Earlier work this paper cites.
An Efficient Memory-Based Morphosyntactic Tagger and Parser for Dutch
Antal van den Bosch, Bertjan Busser, Sander Canisius, and Walter Daelemans. 2007 · 2007
Earlier work this paper cites.
Language-Agnostic BERT Sentence Embedding
Fangxiaoyu Feng, Yinfei Yang, Daniel Cer, Naveen Arivazhagan, and Wei Wang. 2020 · 2007
Earlier work this paper cites.
Nearest Neighbors in High-Dimensional Data: The Emergence and Influence of Hubs
Miloš Radovanović, Alexandros Nanopoulos, and Mirjana Ivanović. 2009 · 2009
Earlier work this paper cites.
Efficient Processing of k Nearest Neighbor Joins Using MapReduce
Wei Lu, Yanyan Shen, Su Chen, and Beng Chin Ooi. 2012 · 2012
Earlier work this paper cites.
A Deep Relevance Matching Model for Ad-Hoc Retrieval
Jiafeng Guo, Yixing Fan, Qingyao Ai, and W. Bruce Croft. 2016 · 2016
Earlier work this paper cites.
Are You a Racist or Am I Seeing Things? Annotator Influence on Hate Speech Detection on Twitter
Zeerak Waseem. 2016 · 2016
Earlier work this paper cites.
Bag of Tricks for Efficient Text Classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov. 2017 · 2017
Earlier work this paper cites.
Learning to Remember Rare Events
Lukasz Kaiser, Ofir Nachum, Aurko Roy, and Samy Bengio. 2017 · 2017
Earlier work this paper cites.
A Structured Self-Attentive Sentence Embedding
Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Deeper Attention to Abusive User Content Moderation
John Pavlopoulos, Prodromos Malakasiotis, and Ion Androutsopoulos. 2017 · 2017
Earlier work this paper cites.
A survey on hate speech detection using natural language processing
Anna Schmidt and Michael Wiegand. 2017 · 2017
Earlier work this paper cites.
Ex Machina: Personal Attacks Seen at Scale
Ellery Wulczyn, Nithum Thain, and Lucas Dixon. 2017 · 2017
Earlier work this paper cites.
Overview of the Second BUCC Shared Task: Spotting Parallel Sentences in Comparable Corpora
Pierre Zweigenbaum, Serge Sharoff, and Reinhard Rapp. 2017 · 2017
Earlier work this paper cites.
Deep Learning for Detecting Cyberbullying Across Multiple Social Media Platforms
Sweta Agrawal and Amit Awekar. 2018 · 2018
Earlier work this paper cites.
The Effects of User Features on Twitter Hate Speech Detection
Elise Fehn Unsvåg and Björn Gambäck. 2018 · 2018
Earlier work this paper cites.
Convolutional Neural Networks for Toxic Comment Classification
Spiros V. Georgakopoulos, Sotiris K. Tasoulis, Aristidis G. Vrahatis, and Vassilis P. Plagianakos. 2018 · 2018
Earlier work this paper cites.
Challenges in discriminating profanity from hate speech
Shervin Malmasi and Marcos Zampieri. 2018 · 2018
Earlier work this paper cites.
Detecting Offensive Tweets in Hindi-English Code-Switched Language
Puneet Mathur, Rajiv Ratn Shah, Ramit Sawhney, and Debanjan Mahata. 2018 · 2018
Earlier work this paper cites.
Will it Blend? Blending Weak and Strong Labeled Data in a Neural Network for Argumentation Mining
Eyal Shnarch, Carlos Alzate, Lena Dankin, Martin Gleize, Yufang Hou, Leshem Choshen, Ranit Aharonov, and Noam Slonim. 2018 · 2018
Earlier work this paper cites.
Identifying Aggression and Toxicity in Comments using Capsule Network
Saurabh Srivastava, Prerna Khurana, and Vartika Tewari. 2018 · 2018
Earlier work this paper cites.
Interpreting Neural Networks with Nearest Neighbors
Eric Wallace, Shi Feng, and Jordan Boyd-Graber. 2018 · 2018
Cited alongside, same era.
Hate Speech Detection is Not as Easy as You May Think: A Closer Look at Model Validation
Aymé Arango, Jorge Pérez, and Barbara Poblete. 2019 · 2019
Cited alongside, same era.
Massively Multilingual Sentence Embeddings for Zero-Shot Cross-Lingual Transfer and Beyond
Mikel Artetxe and Holger Schwenk. 2019 · 2019
Cited alongside, same era.
Stereotypical Bias Removal for Hate Speech Detection Task Using Knowledge-Based Generalizations
Pinkesh Badjatiya, Manish Gupta, and Vasudeva Varma. 2019 · 2019
Cited alongside, same era.
SemEval-2019 Task 5: Multilingual Detection of Hate Speech Against Immigrants and Women in Twitter
Valerio Basile, Cristina Bosco, Elisabetta Fersini, Debora Nozza, Viviana Patti, Francisco Manuel Rangel Pardo, Paolo Rosso, and Manuela Sanguinetti. 2019 · 2019
Cited alongside, same era.
XHate-999: Analyzing and Detecting Abusive Language Across Domains and Languages
Goran Glavaš, Mladen Karan, and Ivan Vulić. 2020 · 2020
Later among the works it cites.
Online Harms White Paper
UK Government. 2020 · 2020
Later among the works it cites.
Retrieval Augmented Language Model Pre-Training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Mingwei Chang. 2020 · 2020
Later among the works it cites.
BERT-kNN: Adding a kNN Search Component to Pretrained Language Models for Better QA
Nora Kassner and Hinrich Schütze. 2020 · 2020
Later among the works it cites.
Generalization through Memorization: Nearest Neighbor Language Models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis. 2020 · 2020
Later among the works it cites.
Toxic Language Detection in Social Media for Brazilian Portuguese: New Dataset and Multilingual Analysis
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning Cross-Lingual Sentence Representations via a Multi-task Dual-Encoder Model
Muthu Chidambaram, Yinfei Yang, Daniel Cer, Steve Yuan, Yunhsuan Sung, Brian Strope, and Ray Kurzweil. 2019 · 2019
Cited alongside, same era.
Cross-lingual Language Model Pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Cited alongside, same era.
Racial Bias in Hate Speech and Abusive Language Detection Datasets
Thomas Davidson, Debasmita Bhattacharya, and Ingmar Weber. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
How Contextual are Contextualized Word Representations? Comparing the Geometry of BERT, ELMo, and GPT-2 Embeddings
Kawin Ethayarajh. 2019 · 2019
Cited alongside, same era.
A Just and Comprehensive Strategy for Using NLP to Address Online Abuse
David Jürgens, Libby Hemphill, and Eshwar Chandrasekharan. 2019 · 2019
Cited alongside, same era.
Hate speech detection: Challenges and solutions
Sean MacAvaney, Hao-Ren Yao, Eugene Yang, Katina Russell, Nazli Goharian, and Ophir Frieder. 2019 · 2019
Cited alongside, same era.
João Augusto Leite, Diego Silva, Kalina Bontcheva, and Carolina Scarton. 2020 · 2020
Later among the works it cites.
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. 2020 · 2020
Later among the works it cites.
Multilingual Offensive Language Identification with Cross-lingual Embeddings
Tharindu Ranasinghe and Marcos Zampieri. 2020 · 2020
Later among the works it cites.
Making Monolingual Sentence Embeddings Multilingual using Knowledge Distillation
Nils Reimers and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers
Wenhui Wang, Furu Wei, Li Dong, Hangbo Bao, Nan Yang, and Ming Zhou. 2020 · 2020
Later among the works it cites.
SemEval-2020 task 12: Multilingual offensive language identification in social media (OffensEval 2020)
Marcos Zampieri, Preslav Nakov, Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Hamdy Mubarak, Leon Derczynski, Zeses Pitenis, and Çağrı Çöltekin. 2020 · 2020
Later among the works it cites.
Detecting Propaganda Techniques in Memes
Dimitar Dimitrov, Bishr Bin Ali, Shaden Shaar, Firoj Alam, Fabrizio Silvestri, Hamed Firooz, Preslav Nakov, and Giovanni Da San Martino. 2021 · 2021
Closest in time.
Augmenting Transformers with KNN-Based Composite Memory for Dialog
Angela Fan, Claire Gardent, Chloé Braud, and Antoine Bordes. 2021 · 2021
Closest in time.
Toxic Comment Classification Challenge
Jigsaw. 2018 · 2021
Closest in time.
Jigsaw Multilingual Toxic Comment Classification
Jigsaw Multilingual. 2020 · 2021
Closest in time.
Billion-Scale Similarity Search with GPUs
Jeff Johnson, Matthijs Douze, and Hervé Jégou. 2021 · 2021
Closest in time.
Detecting Abusive Language on Online Platforms: A Critical Analysis
Preslav Nakov, Vibha Nayak, Kyle Dent, Ameya Bhatawdekar, Sheikh Muhammad Sarwar, Momchil Hardalov, Yoan Dinkov, Dimitrina Zlatkova, Guillaume Bouchard, and Isabelle Augenstein. 2021 · 2021
Closest in time.
Detecting harmful memes and their targets
Shraman Pramanick, Dimitar Dimitrov, Rituparna Mukherjee, Shivam Sharma, Md. Shad Akhtar, Preslav Nakov, and Tanmoy Chakraborty. 2021a · 2021
Closest in time.
MOMENTA: A multimodal framework for detecting harmful memes and their targets
Shraman Pramanick, Shivam Sharma, Dimitar Dimitrov, Md. Shad Akhtar, Preslav Nakov, and Tanmoy Chakraborty. 2021b · 2021
Closest in time.
SOLID: A large-scale semi-supervised dataset for offensive language identification
Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Marcos Zampieri, and Preslav Nakov. 2021 · 2021
Closest in time.
Offensive language identification in Dravidian code mixed social media text
Sunil Saumya, Abhinav Kumar, and Jyoti Prakash Singh. 2021 · 2021
Closest in time.
Inducing Language-Agnostic Multilingual Representations
Wei Zhao, Steffen Eger, Johannes Bjerva, and Isabelle Augenstein. 2021 · 2021
Closest in time.