Fetching the paper…
Reading the bibliography…
Existing backdoor defense methods are only effective for limited trigger types.
Weight poisoning attacks on pre-trained models
Keita Kurita, Paul Michel, and Graham Neubig. 2020 · 2004
Earlier work this paper cites.
Badnl: Backdoor attacks against NLP models
Xiaoyi Chen, Ahmed Salem, Michael Backes, Shiqing Ma, and Yang Zhang. 2020 · 2006
Earlier work this paper cites.
Yiming Li, Baoyuan Wu, Yong Jiang, Zhifeng Li, and Shu-Tao Xia. 2020 · 2007
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew Maas, Raymond E Daly, Peter T Pham, Dan Huang, Andrew Y Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D Manning, Andrew Y Ng, and Christopher Potts. 2013 · 2013
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Earlier work this paper cites.
Targeted backdoor attacks on deep learning systems using data poisoning
Xinyun Chen, Chang Liu, Bo Li, Kimberly Lu, and Dawn Song. 2017 · 2017
Earlier work this paper cites.
Neural trojans
Yuntao Liu, Yang Xie, and Ankur Srivastava. 2017 · 2017
Earlier work this paper cites.
Weakly-supervised neural text classification
Yu Meng, Jiaming Shen, Chao Zhang, and Jiawei Han. 2018 · 2018
Earlier work this paper cites.
Poison frogs! targeted clean-label poisoning attacks on neural networks
Ali Shafahi, W. Ronny Huang, Mahyar Najibi, Octavian Suciu, Christoph Studer, Tudor Dumitras, and Tom Goldstein. 2018 · 2018
Cited alongside, same era.
A backdoor attack against lstm-based text classification systems
Jiazhu Dai, Chuanshuai Chen, and Yufeng Li. 2019 · 2019
Cited alongside, same era.
Badnets: Evaluating backdooring attacks on deep neural networks
Tianyu Gu, Kang Liu, Brendan Dolan-Gavitt, and Siddharth Garg. 2019 · 2019
Cited alongside, same era.
Crossweigh: Training named entity tagger from imperfect annotations
Zihan Wang, Jingbo Shang, Liyuan Liu, Lihao Lu, Jiacheng Liu, and Jiawei Han. 2019 · 2019
Cited alongside, same era.
ELECTRA: pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020 · 2020
Cited alongside, same era.
Contextualized weak supervision for text classification
Mitigating backdoor attacks in lstm-based text classification systems by backdoor keyword identification
Chuanshuai Chen and Jiazhu Dai. 2021 · 2021
Later among the works it cites.
Textual backdoor attacks can be more harmful via two simple tricks
Yangyi Chen, Fanchao Qi, Zhiyuan Liu, and Maosong Sun. 2021 · 2021
Later among the works it cites.
Triggerless backdoor attack for NLP tasks with clean labels
Leilei Gan, Jiwei Li, Tianwei Zhang, Xiaoya Li, Yuxian Meng, Fei Wu, Shangwei Guo, and Chun Fan. 2021 · 2021
Later among the works it cites.
Bfclass: A backdoor-free text classification framework
Zichao Li, Dheeraj Mekala, Chengyu Dong, and Jingbo Shang. 2021 · 2021
Later among the works it cites.
ONION: A simple and effective defense against textual backdoor attacks
Fanchao Qi, Yangyi Chen, Mukai Li, Yuan Yao, Zhiyuan Liu, and Maosong Sun. 2021a · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dheeraj Mekala and Jingbo Shang. 2020 · 2020
Cited alongside, same era.
Text classification using label names only: A language model self-training approach
Yu Meng, Yunyi Zhang, Jiaxin Huang, Chenyan Xiong, Heng Ji, Chao Zhang, and Jiawei Han. 2020 · 2020
Cited alongside, same era.
Clean-label backdoor attacks on video recognition models
Shihao Zhao, Xingjun Ma, Xiang Zheng, James Bailey, Jingjing Chen, and Yu-Gang Jiang. 2020 · 2020
Cited alongside, same era.
Backdoor embedding in convolutional neural network models via invisible perturbation
Haoti Zhong, Cong Liao, Anna Cinzia Squicciarini, Sencun Zhu, and David J. Miller. 2020 · 2020
Cited alongside, same era.
Hidden killer: Invisible textual backdoor attacks with syntactic trigger
Fanchao Qi, Mukai Li, Yangyi Chen, Zhengyan Zhang, Zhiyuan Liu, Yasheng Wang, and Maosong Sun. 2021b · 2021
Later among the works it cites.
"average" approximates "first principal component"? an empirical analysis on representations from neural language models
Zihan Wang, Chengyu Dong, and Jingbo Shang. 2021a · 2021
Later among the works it cites.
X-class: Text classification with extremely weak supervision
Zihan Wang, Dheeraj Mekala, and Jingbo Shang. 2021b · 2021
Later among the works it cites.