T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” in Advances in neural information processing systems (NeurIPS) , vol. 33. Curran Associates, Inc., 2020, pp. 1877–1901
1901
Earlier work this paper cites.
N. Carlini, S. Chien, M. Nasr, S. Song, A. Terzis, and F. Tramer, “Membership inference attacks from first principles,” in 43rd IEEE Symposium on Security and Privacy (S&P) . IEEE, 2022, pp. 1897–1914
1914
Earlier work this paper cites.
P. Samarati and L. Sweeney, “Protecting privacy when disclosing information: k k -anonymity and its enforcement through generalization and suppression,” Computer Science Laboratory, SRI International, Tech. Rep., 1998. [Online]. Available: http://www.csl.sri.com/papers/sritr-98-04/
1998
Earlier work this paper cites.
B. Klimt and Y. Yang, “Introducing the Enron corpus,” in International Conference on Email and Anti-Spam, CEAS , vol. 45, 2004, pp. 92–96
2004
Earlier work this paper cites.
C. Dwork, K. Kenthapadi, F. McSherry, I. Mironov, and M. Naor, “Our data, ourselves: Privacy via distributed noise generation,” in Advances in Cryptology – EUROCRYPT 2006 . Springer, 2006, pp. 486–503
2006
Earlier work this paper cites.
C. Dwork, F. McSherry, K. Nissim, and A. Smith, “Calibrating noise to sensitivity in private data analysis,” in Theory of cryptography conference, TCC . Springer, 2006, pp. 265–284
2006
Earlier work this paper cites.
P. Golle, “Revisiting the uniqueness of simple demographics in the us population,” in 5th ACM Workshop on Privacy in Electronic Society . ACM, 2006, pp. 77–80
2006
Earlier work this paper cites.
E. Hovy, M. Marcus, M. Palmer, L. Ramshaw, and R. Weischedel, “OntoNotes: The 90% solution,” in Human Language Technology Conference of the NAACL . ACL, 2006, pp. 57–60
2006
Earlier work this paper cites.
D. Nadeau and S. Sekine, “A survey of named entity recognition and classification,” Lingvisticae Investigationes , vol. 30, no. 1, pp. 3–26, 2007
2007
Earlier work this paper cites.
S. Bird, E. Klein, and E. Loper, Natural language processing with Python: analyzing text with the Natural Language Toolkit . O’Reilly Media, 2009
2009
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” arXiv preprint arXiv:1508.07909 [cs.CL] , 2015
Original
2015
Earlier work this paper cites.
M. Abadi, A. Chu, I. Goodfellow, H. B. McMahan, I. Mironov, K. Talwar, and L. Zhang, “Deep learning with differential privacy,” in 23rd ACM SIGSAC Conference on Computer and Communications Security, CCS . ACM, 2016, pp. 308–318
2016
Earlier work this paper cites.
European Union, “Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 april 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46/EC (General Data Protection Regulation),” Official Journal , vol. L 110, pp. 1–88, 2016
2016
Earlier work this paper cites.
G. Lample, M. Ballesteros, S. Subramanian, K. Kawakami, and C. Dyer, “Neural architectures for named entity recognition,” in Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT . ACL, 2016, pp. 260–270
2016
Earlier work this paper cites.
R. Shokri, M. Stronati, C. Song, and V. Shmatikov, “Membership inference attacks against machine learning models,” in 37th IEEE symposium on security and privacy (S&P) . IEEE, 2017, pp. 3–18
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems (NeurIPS) , vol. 30, p. 6000–6010, 2017
2017
Earlier work this paper cites.
A. Fan, M. Lewis, and Y. Dauphin, “Hierarchical neural story generation,” in 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . ACL, 2018, pp. 889–898
2018
Earlier work this paper cites.
A. Akbik, T. Bergmann, D. Blythe, K. Rasul, S. Schweter, and R. Vollgraf, “FLAIR: An easy-to-use framework for state-of-the-art NLP,” in 2019 Conference of the North American Chapter of the ACL (demonstrations) . ACL, 2019, pp. 54–59
2019
Earlier work this paper cites.
P. Budzianowski and I. Vulić, “Hello, it’s GPT-2 – How can I help you? towards the use of pretrained language models for task-oriented dialogue systems,” arXiv preprint arXiv:1907.05774 [cs.CL] , 2019
Original
2019
Earlier work this paper cites.
N. Carlini, C. Liu, Ú. Erlingsson, J. Kos, and D. Song, “The secret sharer: Evaluating and testing unintended memorization in neural networks,” in 28th USENIX Security Symposium (USENIX Security 19) , 2019, pp. 267–284
2019
Earlier work this paper cites.
I. Chalkidis, I. Androutsopoulos, and N. Aletras, “Neural legal judgment prediction in English,” in 57th Annual Meeting of the Association for Computational Linguistics, ACL . ACL, 2019, pp. 4317–4323
2019
Earlier work this paper cites.
M. X. Chen, B. N. Lee, G. Bansal, Y. Cao, S. Zhang, J. Lu, J. Tsay, Y. Wang, A. M. Dai, Z. Chen et al. , “Gmail smart compose: Real-time assisted writing,” in 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , 2019, pp. 2287–2295
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT . ACL, 2019, pp. 4171–4186
2019
Earlier work this paper cites.