Fetching the paper…
Reading the bibliography…
Deep Neural Network (DNN) classifiers are known to be vulnerable to Trojan or backdoor attacks, where the classifier is manipulated such that it misclassifies any input containing an attacker-determined Trojan trigger.
LIII. On lines and planes of closest fit to systems of points in space
Karl Pearson · 1901
Earlier work this paper cites.
Analysis of a Complex of Statistical Variables into Principal Components
Harold Hotelling · 1933
Earlier work this paper cites.
A Density-Based Algorithm for Discovering Clusters a Density-Based Algorithm for Discovering Clusters in Large Spatial Databases with Noise
Martin Ester, Hans-Peter Kriegel, Jörg Sander, and Xiaowei Xu · 1996
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Seeing Stars: Exploiting Class Relationships for Sentiment Categorization with Respect to Rating Scales
Bo Pang and Lillian Lee · 2005
Earlier work this paper cites.
Spam Filtering with Naive Bayes-which Naive Bayes?
Vangelis Metsis, Ion Androutsopoulos, and Georgios Paliouras · 2006
Earlier work this paper cites.
Learning Word Vectors for Sentiment Analysis
Andrew Maas, Raymond E Daly, Peter T Pham, Dan Huang, Andrew Y Ng, and Christopher Potts · 2011
Earlier work this paper cites.
Recursive Deep Models for Semantic Compositionality over a Sentiment Treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D Manning, Andrew Y Ng, and Christopher Potts · 2013
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Caffe: Convolutional Architecture for Fast Feature Embedding
Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross Girshick, Sergio Guadarrama, and Trevor Darrell · 2014
Earlier work this paper cites.
Convolutional Neural Networks for Sentence Classification
Yoon Kim · 2014
Earlier work this paper cites.
Intriguing Properties of Neural Networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
How Transferable are Features in Deep Neural Networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson · 2014
Earlier work this paper cites.
Implementing a CNN for Text Classification in Tensorflow
Denny Britz · 2015
Earlier work this paper cites.
Collective Opinion Spam Detection: Bridging Review Networks and Metadata
Shebuti Rayana and Leman Akoglu · 2015
Earlier work this paper cites.
Business Reviews Classification Using Sentiment Analysis
Andreea Salinca · 2015
Earlier work this paper cites.
Transfer Learning for Speech and Language Processing
Dong Wang and Thomas Fang Zheng · 2015
Earlier work this paper cites.
Character-level Convolutional Networks for Text Classification
Xiang Zhang, Junbo Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Deep Learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Earlier work this paper cites.
Crafting Adversarial Input Sequences for Recurrent Neural Networks
Nicolas Papernot, Patrick McDaniel, Ananthram Swami, and Richard Harang · 2016
Earlier work this paper cites.
Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter
Zeerak Waseem and Dirk Hovy · 2016
Earlier work this paper cites.
Towards Evaluating the Robustness of Neural Networks
Nicholas Carlini and David Wagner · 2017
Cited alongside, same era.
Automated Hate Speech Detection and the Problem of Offensive Language
Thomas Davidson, Dana Warmsley, Michael Macy, and Ingmar Weber · 2017
Cited alongside, same era.
Badnets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain
Tianyu Gu, Brendan Dolan-Gavitt, and Siddharth Garg · 2017
Cited alongside, same era.
Toward Controlled Generation of Text
Zhiting Hu, Zichao Yang, Xiaodan Liang, Ruslan Salakhutdinov, and Eric P Xing · 2017
Cited alongside, same era.
Trojaning Attack on Neural Networks
Yingqi Liu, Shiqing Ma, Yousra Aafer, Wen-Chuan Lee, Juan Zhai, Weihang Wang, and Xiangyu Zhang · 2017
Cited alongside, same era.
Towards Deep Learning Models Resistant to Adversarial Attacks
Provably Robust Boosted Decision Stumps and Trees against Adversarial Attacks
Maksym Andriushchenko and Matthias Hein · 2019
Later among the works it cites.
Deepinspect: A Black-Box Trojan Detection and Mitigation Framework for Deep Neural Networks
Huili Chen, Cheng Fu, Jishen Zhao, and Farinaz Koushanfar · 2019
Later among the works it cites.
A Backdoor Attack against LSTM-based Text Classification Systems
Jiazhu Dai, Chuanshuai Chen, and Yufeng Li · 2019
Later among the works it cites.
STRIP: A Defence against Trojan Attacks on Deep Neural Networks
Yansong Gao, Change Xu, Derui Wang, Shiping Chen, Damith C Ranasinghe, and Surya Nepal · 2019
Later among the works it cites.
BadNets: Evaluating Backdooring Attacks on Deep Neural Networks
Tianyu Gu, Kang Liu, Brendan Dolan-Gavitt, and Siddharth Garg · 2019
Later among the works it cites.
TABOR: A Highly Accurate Approach to Inspecting and Restoring Trojan Backdoors in AI Systems
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt, Dimitris Tsipras, and Adrian Vladu · 2017
Cited alongside, same era.
Universal Adversarial Perturbations
Seyed-Mohsen Moosavi-Dezfooli, Alhussein Fawzi, Omar Fawzi, and Pascal Frossard · 2017
Cited alongside, same era.
Fast Feature Fool: A Data Independent Approach to Universal Adversarial Perturbations
Konda Reddy Mopuri, Utsav Garg, and R Venkatesh Babu · 2017
Cited alongside, same era.
Compass: Spatio Temporal Sentiment Analysis of US Election What Twitter Says!
Debjyoti Paul, Feifei Li, Murali Krishna Teja, Xin Yu, and Richie Frost · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Generating Natural Language Adversarial Examples
Moustafa Alzantot, Yash Sharma, Ahmed Elgohary, Bo-Jhang Ho, Mani Srivastava, and Kai-Wei Chang · 2018
Cited alongside, same era.
Detecting Backdoor Attacks on Deep Neural Networks by Activation Clustering
Bryant Chen, Wilka Carvalho, Nathalie Baracaldo, Heiko Ludwig, Benjamin Edwards, Taesung Lee, Ian Molloy, and Biplav Srivastava · 2018
Cited alongside, same era.
Wenbo Guo, Lun Wang, Xinyu Xing, Min Du, and Dawn Song · 2019
Later among the works it cites.
TEXTBUGGER: Generating Adversarial Text against Real-world Applications
Jinfeng Li, Shouling Ji, Tianyu Du, Bo Li, and Ting Wang · 2019
Later among the works it cites.
Generating Natural Language Adversarial Examples Through Probability Weighted Word Saliency
Shuhuai Ren, Yihe Deng, Kun He, and Wanxiang Che · 2019
Later among the works it cites.
Towards the First Adversarially Robust Neural Network Model on MNIST
L Schott, J Rauber, M Bethge, and W Brendel · 2019
Later among the works it cites.
Spam Review Detection Using Deep Learning
G. M. Shahariar, Swapnil Biswas, Faiza Omar, Faisal Muhammad Shah, and Samiha Binte Hassan · 2019
Later among the works it cites.
Universal Adversarial Triggers for Attacking and Analyzing NLP
Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh · 2019
Later among the works it cites.
Neural Cleanse: Identifying and Mitigating Backdoor Attacks in Neural Networks
Bolun Wang, Yuanshun Yao, Shawn Shan, Huiying Li, Bimal Viswanath, Haitao Zheng, and Ben Y Zhao · 2019
Later among the works it cites.
Latent Backdoor Attacks on Deep Neural Networks
Yuanshun Yao, Huiying Li, Haitao Zheng, and Ben Y Zhao · 2019
Later among the works it cites.
How to Backdoor Federated Learning
Eugene Bagdasaryan, Andreas Veit, Yiqing Hua, Deborah Estrin, and Vitaly Shmatikov · 2020
Later among the works it cites.
Bae: Bert-based Adversarial Examples for Text Classification
Siddhant Garg and Goutham Ramakrishnan · 2020
Later among the works it cites.
Fake Consumer Review Detection Using Deep Neural Networks Integrating Word Embeddings and Emotion Mining
Petr Hajek, Aliaksandr Barushka, and Michal Munk · 2020
Later among the works it cites.
Is Bert Really Robust? A Strong Baseline for Natural Language Attack on Text Classification and Entailment
Di Jin, Zhijing Jin, Joey Tianyi Zhou, and Peter Szolovits · 2020
Later among the works it cites.
BERT-ATTACK: Adversarial Attack against BERT Using BERT
Linyang Li, Ruotian Ma, Qipeng Guo, Xiangyang Xue, and Xipeng Qiu · 2020
Later among the works it cites.
Fakeddit: A New Multimodal Benchmark Dataset for Fine-grained Fake News Detection
Kai Nakamura, Sharon Levy, and William Yang Wang · 2020
Later among the works it cites.
Backdoor Attacks against Transfer Learning with Pre-trained Deep Learning Models
Shuo Wang, Surya Nepal, Carsten Rudolph, Marthie Grobler, Shangyu Chen, and Tianle Chen · 2020
Later among the works it cites.
Fast-UAP: An Algorithm for Expediting Universal Adversarial Perturbation Generation Using the Orientations of Perturbation Vectors
Jiazhu Dai and Le Shu · 2021
Closest in time.