Defending against machine learning model stealing attacks using deceptive perturbations
Taesung Lee, Benjamin Edwards, Ian Molloy, and Dong Su. 2019 · 2019
Later among the works it cites.
On evaluation of adversarial perturbations for sequence-to-sequence models
Paul Michel, Xian Li, Graham Neubig, and Juan Miguel Pino. 2019 · 2019
Later among the works it cites.
Knockoff nets: Stealing functionality of black-box models
Tribhuvanesh Orekondy, Bernt Schiele, and Mario Fritz. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
Evaluating gender bias in machine translation
Gabriel Stanovsky, Noah A Smith, and Luke Zettlemoyer. 2019 · 2019
Later among the works it cites.
Insertion transformer: Flexible sequence generation via insertion operations
Mitchell Stern, William Chan, Jamie Kiros, and Jakob Uszkoreit. 2019 · 2019
Later among the works it cites.
Multilingual neural machine translation with knowledge distillation
Xu Tan, Yi Ren, Di He, Tao Qin, and Tie-Yan Liu. 2019 · 2019
Later among the works it cites.
Universal adversarial triggers for attacking and analyzing NLP
Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh. 2019 · 2019
Later among the works it cites.
Knowing when to stop: Evaluation and verification of conformity to output-size specifications
Chenglong Wang, Rudy Bunel, Krishnamurthy Dvijotham, Po-Sen Huang, Edward Grefenstette, and Pushmeet Kohli. 2019 · 2019
Later among the works it cites.
Understanding knowledge distillation in non-autoregressive machine translation
Original
Chunting Zhou, Graham Neubig, and Jiatao Gu. 2020 · 2019
Later among the works it cites.
Exploring connections between active learning and model extraction
Varun Chandrasekaran, K Chaudhari, Irene Giacomelli, Somesh Jha, and Songbai Yan. 2020 · 2020
Closest in time.
Seq2Sick: Evaluating the robustness of sequence-to-sequence models with adversarial examples
Minhao Cheng, Jinfeng Yi, Huan Zhang, Pin-Yu Chen, and Cho-Jui Hsieh. 2020 · 2020
Closest in time.
Membership inference attacks on sequence-to-sequence models: Is my data in your machine translation system?
Sorami Hisamoto, Matt Post, and Kevin Duh. 2020 · 2020
Closest in time.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Closest in time.
A scalable approach to reducing gender bias in Google Translate
Melvin Johnson. 2020 · 2020
Closest in time.
Thieves on sesame street! Model extraction of BERT-based APIs
Kalpesh Krishna, Gaurav Singh Tomar, Ankur P Parikh, Nicolas Papernot, and Mohit Iyyer. 2020 · 2020
Closest in time.
Self-distillation amplifies regularization in Hilbert space
Hossein Mobahi, Mehrdad Farajtabar, and Peter L Bartlett. 2020 · 2020
Closest in time.
Prediction poisoning: Towards defenses against dnn model stealing attacks
Tribhuvanesh Orekondy, Bernt Schiele, and Mario Fritz. 2020 · 2020
Closest in time.