Interpretable adversarial training for text
Samuel Barham and Soheil Feizi · 2019
Later among the works it cites.
Evaluating and enhancing the robustness of dialogue systems: A case study on a negotiation agent
Minhao Cheng, Wei Wei, and Cho-Jui Hsieh · 2019
Later among the works it cites.
Certified adversarial robustness via randomized smoothing
Jeremy Cohen, Elan Rosenfeld, and Zico Kolter · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Later among the works it cites.
Defense against adversarial images using web-scale nearest-neighbor search
Abhimanyu Dubey, Laurens van der Maaten, Zeki Yalniz, Yixuan Li, and Dhruv Mahajan · 2019
Later among the works it cites.
Achieving verified robustness to symbol substitutions via interval bound propagation
Po-Sen Huang, Robert Stanforth, Johannes Welbl, Chris Dyer, Dani Yogatama, Sven Gowal, Krishnamurthy Dvijotham, and Pushmeet Kohli · 2019
Later among the works it cites.
Certified robustness to adversarial word substitutions
Robin Jia, Aditi Raghunathan, Kerem Göksel, and Percy Liang · 2019
Later among the works it cites.
Certified robustness to adversarial examples with differential privacy
Mathias Lecuyer, Vaggelis Atlidakis, Roxana Geambasu, Daniel Hsu, and Suman Jana · 2019
Later among the works it cites.
On evaluation of adversarial perturbations for sequence-to-sequence models
Paul Michel, Xian Li, Graham Neubig, and Juan Miguel Pino · 2019
Later among the works it cites.
Generating natural language adversarial examples through probability weighted word saliency
Shuhuai Ren, Yihe Deng, Kun He, and Wanxiang Che · 2019
Later among the works it cites.
Interpretable adversarial perturbation in input embedding space for text
Motoki Sato, Jun Suzuki, Shindo, and Yuji Matsumoto · 2019
Later among the works it cites.
FreeLB: Enhanced adversarial training for language understanding
Chen Zhu, Yu Cheng, Zhe Gan, Siqi Sun, Tom Goldstein, and Jingjing Liu · 2019
Later among the works it cites.
Robustness verification for transformers
Zhouxing Shi, Kai-Wei Chang Huan Zhang, Minlie Huang, and Cho-Jui Hsieh · 2020
Closest in time.
Towards stable and efficient training of verifiably robust neural networks
Huan Zhang, Hongge Chen, Chaowei Xiao, Bo Li, Duane Boning, and Cho-Jui Hsieh · 2020
Closest in time.
Evaluating and enhancing the robustness of neural network-based dependency parsing models with adversarial examples
Xiaoqing Zheng, Jiehang Zeng, Yi Zhou, Cho-Jui Hsieh, Minhao Cheng, and Xuanjing Huang · 2020
Closest in time.