Fetching the paper…
Reading the bibliography…
Large scale adoption of large language models has introduced a new era of convenient knowledge transfer for a slew of natural language processing tasks.
Language Models are Few-Shot Learners
Brown, Tom, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
Stolen memories: Leveraging model memorization for calibrated white-box membership inference
Leino, Klas and Matt Fredrikson. 2020 · 1906
Earlier work this paper cites.
Privacy preserving text representation learning
Beigi, Ghazaleh, Kai Shu, Ruocheng Guo, Suhang Wang, and Huan Liu. 2019 · 1907
Earlier work this paper cites.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Liu, Yinhan, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Local Differential Privacy for Deep Learning
Mahawaga Arachchige, Pathum Chamikara, Peter Bertok, Ibrahim Khalil, Dongxi Liu, Seyit Camtepe, and Mohammed Atiquzzaman. 2020 · 1908
Earlier work this paper cites.
Privacy- And utility-preserving textual analysis via calibrated multivariate perturbations
Feyisetan, Oluwaseyi, Borja Balle, Thomas Drake, and Tom Diethe. 2020 · 1910
Earlier work this paper cites.
Deep Poisoning Functions: Towards Robust Privacy-safe Image Data Sharing
Guo, Hao, Brian Dolhansky, Eric Hsin, Phong Dinh, Song Wang, and Cristian Canton Ferrer. 2019 · 1912
Earlier work this paper cites.
Multilingual is not enough: BERT for Finnish
Virtanen, Antti, Jenna Kanerva, Rami Ilo, Jouni Luoma, Juhani Luotolahti, Tapio Salakoski, Filip Ginter, and Sampo Pyysalo. 2019 · 1912
Earlier work this paper cites.
The Effects of Corpus Size and Homogeneity on Language Model Quality
Rose, Tony G. 1997 · 1997
Earlier work this paper cites.
Can Pseudonymity Really Guarantee Privacy?
Rao, Josyula, Josyula Rao, and Pankaj Rohatgi. 2000 · 2000
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the NLP world
Joshi, Pratik, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2004
Earlier work this paper cites.
Information Leakage in Embedding Models
Song, Congzheng and Ananth Raghunathan. 2020 · 2004
Earlier work this paper cites.
We Need to Talk About Random Splits
Søgaard, Anders, Sebastian Ebert, Jasmijn Bastings, and Katja Filippova. 2021 · 2005
Earlier work this paper cites.
Calibrating noise to sensitivity in private data analysis
Dwork, Cynthia, Frank McSherry, Kobbi Nissim, and Adam Smith. 2006 · 2006
Earlier work this paper cites.
You are what you say: privacy risks of public mentions
Frankowski, Dan, Dan Cosley, Shilad Sen, Loren Terveen, and John Riedl. 2006 · 2006
Earlier work this paper cites.
How To Break Anonymity of the Netflix Prize Dataset
Narayanan, Arvind and Vitaly Shmatikov. 2006 · 2006
Earlier work this paper cites.
The boundary between privacy and utility in data publishing
Rastogi, Vibhor, Dan Suciu, and Sungho Hong. 2007 · 2007
Earlier work this paper cites.
Representation of the Sexes in Language
Stahlberg, Dagmar, Friederike Braun, Lisa Irmen, and Sabine Sczesny. 2007 · 2007
Earlier work this paper cites.
The cost of privacy: destruction of data-mining utility in anonymized data publishing
Brickell, Justin and Vitaly Shmatikov. 2008 · 2008
Earlier work this paper cites.
Visualizing Data using t-SNE
Maaten, Laurens van der and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Robust de-anonymization of large sparse datasets
Narayanan, Arvind and Vitaly Shmatikov. 2008 · 2008
Earlier work this paper cites.
Anonymizing transaction databases for publication
Xu, Yabo, Ke Wang, Ada Wai Chee Fu, and Philip S. Yu. 2008 · 2008
Earlier work this paper cites.
On the tradeoff between privacy and utility in data publishing
Li, Tiancheng and Ninghui Li. 2009 · 2009
Earlier work this paper cites.
De-anonymizing social networks
Narayanan, Arvind and Vitaly Shmatikov. 2009 · 2009
Earlier work this paper cites.
Anonymity, unobservability, and pseudonymity – A proposal for terminology
Pfitzmann, Andreas and Marit Kohntopp. 2001 · 2009
Earlier work this paper cites.
The Effect of Corpus Size on Case Frame Acquisition for Discourse Analysis
Sasano, Ryohei, Daisuke Kawahara, and Sadao Kurohashi. 2009 · 2009
Earlier work this paper cites.
Privacy Calculus on Social Networking Sites: Explorative Evidence from Germany and USA
Krasnova, Hanna and Natasha F. Veltri. 2010 · 2010
Earlier work this paper cites.
Differentially Private Representation for NLP: Formal Guarantee and An Empirical Study on Privacy and Fairness
Lyu, Lingjuan, Xuanli He, and Yitong Li. 2020 · 2010
Earlier work this paper cites.
A Differentially Private Text Perturbation Method Using Regularized Mahalanobis Metric
Xu, Zekun, Abhinav Aggarwal, Oluwaseyi Feyisetan, and Nathanael Teissier. 2020 · 2010
Earlier work this paper cites.
Smart Metering De-Pseudonymization
Jawurek, Marek, Martin Johns, and Konrad Rieck. 2011 · 2011
Earlier work this paper cites.
Joint Link-Attribute User Identity Resolution in Online Social Networks Categories and Subject Descriptors
Bartunov, Sergey, Anton Korshunov, Seung-taek Park, Wonho Ryu, and Hyungdong Lee. 2012 · 2012
Earlier work this paper cites.
Expanding Parallel Resources for Medium-Density Languages for Free
Iliev, Georgi and Angel Genov. 2012 · 2012
Earlier work this paper cites.
Beware of what you share: Inferring home location in social networks
Pontes, Tatiana, Gabriel Magno, Marisa Vasconcelos, Aditi Gupta, Jussara Almeida, Ponnurangam Kumaraguru, and Virgilio Almeida. 2012 · 2012
Earlier work this paper cites.
On the identity anonymization of high-dimensional rating data
Sun, Xiaoxun, Hua Wang, and Yanchun Zhang. 2012 · 2012
Earlier work this paper cites.
Broadening the Scope of Differential Privacy Using Metrics
Chatzikokolakis, Konstantinos, Miguel E Andrés, Nicolás E Bordenabe, and Catuscia Palamidessi. 2013 · 2013
Earlier work this paper cites.
The algorithmic foundations of differential privacy
Dwork, Cynthia and Aaron Roth. 2013 · 2013
Earlier work this paper cites.
Exploiting innocuous activity for correlating users across sites
Goga, Oana, Howard Lei, Sree Hari Krishnan Parthasarathi, Gerald Friedland, Robin Sommer, and Renata Teixeira. 2013 · 2013
Earlier work this paper cites.
Deanonymisation of clients in bitcoin P2P network
Biryukov, Alex, Dmitry Khovratovich, and Ivan Pustogarov. 2014 · 2014
Earlier work this paper cites.
Echo Chamber or Public Sphere? Predicting Political Orientation and Measuring Political Homophily in Twitter Using Big Data
Colleoni, Elanor, Alessandro Rozza, and Adam Arvidsson. 2014 · 2014
Earlier work this paper cites.
Privacy, anonymity, and big data in the social sciences
Daries, Jon P., Justin Reich, Jim Waldo, Elise M. Young, Jonathan Whittinghill, Andrew Dean Ho, Daniel Thomas Seaton, and Isaac Chuang. 2014 · 2014
Earlier work this paper cites.
RAPPOR: Randomized aggregatable privacy-preserving ordinal response
Erlingsson, Ulfar, Vasyl Pihur, and Aleksandra Korolova. 2014 · 2014
Earlier work this paper cites.
Political Ideology Detection Using Recursive Neural Networks
Iyyer, Mohit, Peter Enns, Jordan Boyd-Graber, and Philip Resnik. 2014 · 2014
Earlier work this paper cites.
Glove: Global Vectors for Word Representation
Pennington, Jeffrey, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Constructing elastic distinguishability metrics for location privacy
Chatzikokolakis, Konstantinos, Catuscia Palamidessi, and Marco Stronati. 2015 · 2015
Earlier work this paper cites.
Model inversion attacks that exploit confidence information and basic countermeasures
Fredrikson, Matt, Somesh Jha, and Thomas Ristenpart. 2015 · 2015
Earlier work this paper cites.
Data, privacy, and the greater good
Horvitz, Eric and Deirdre Mulligan. 2015 · 2015
Earlier work this paper cites.
User Review Sites as a Resource for Large-Scale Sociolinguistic Studies
Hovy, Dirk, Anders Johannsen, and Anders Søgaard. 2015 · 2015
Earlier work this paper cites.
Conservative or liberal? Personalized differential privacy
Jorgensen, Z., T. Yu, and G. Cormode. 2015 · 2015
Cited alongside, same era.
Grammatical Gender in Norwegian: Language Acquisition and Language Change
Rodina, Yulia and Marit Westergaard. 2015 · 2015
Cited alongside, same era.
Privacy-preserving deep learning
Shokri, Reza and Vitaly Shmatikov. 2015 · 2015
Cited alongside, same era.
User tolerance of privacy abuse on mobile Internet and the country level of development
Callanan, Cormac, Borka Jerman-Blažič, and Andrej Jerman Blažič. 2016 · 2016
Cited alongside, same era.
Bucking the Linguistic Binary: Gender Neutral Language in English, Swedish, French, and German
Hord, Levi. 2016 · 2016
Cited alongside, same era.
Discrete distribution estimation under local privacy
Kairouz, Peter, Keith Bonawitz, and Daniel Ramage. 2016 · 2016
To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks
Peters, Matthew E., Sebastian Ruder, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Language Models as Knowledge Bases?
Petroni, Fabio, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Later among the works it cites.
Scalable Differential Privacy with Certified Robustness in Adversarial Learning
Phan, Nhat Hai, My T. Thai, Ruoming Jin, Han Hu, and Dejing Dou. 2019 · 2019
Later among the works it cites.
How Multilingual is Multilingual BERT?
Pires, Telmo, Eva Schlinger, and Dan Garrette. 2019 · 2019
Later among the works it cites.
Is Multilingual BERT Fluent in Language Generation?
Rönnqvist, Samuel, Jenna Kanerva, Tapio Salakoski, and Filip Ginter. 2019 · 2019
Later among the works it cites.
Privacy, Trust and Ethical Issues
Shadbolt, Nigel, Kieron O’Hara, David De Roure, and Wendy Hall. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gendered discourse in German chatroom conversations: the use of modal particles by young adults
Kokovidis, Alexandra. 2015 · 2016
Cited alongside, same era.
Dependency Based Embeddings for Sentence Classification Tasks
Komninos, Alexandros and Suresh Manandhar. 2016 · 2016
Cited alongside, same era.
Thumbs up for privacy?: Differences in online self-disclosure behavior across national cultures
Reed, Philip J., Emma S. Spiro, and Carter T. Butts. 2016 · 2016
Cited alongside, same era.
Social Media Research: A Guide to Ethics
Townsend, Leanne and Claire Wallace. 2016 · 2016
Cited alongside, same era.
Using randomized response for differential privacy preserving data collection
Wang, Yue, Xintao Wu, and Donghui Hu. 2016 · 2016
Cited alongside, same era.
Comparative studies of variation in the use of grammatical gender in the Danish and Dutch DP in the speech of youngsters: Free versus bound morphemes
Cornips, Leonie and Frans Gregersen. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Demystifying Membership Inference Attacks in Machine Learning as a Service
Truex, Stacey, Ling Liu, Mehmet Emre Gursoy, Lei Yu, and Wenqi Wei. 2019 · 2019
Later among the works it cites.
Transferable Attention for Domain Adaptation
Wang, Ximei, Liang Li, Weirui Ye, Mingsheng Long, and Jianmin Wang. 2019 · 2019
Later among the works it cites.
Privacy-aware text rewriting
Xu, Qiongkai, Lizhen Qu, Chenchen Xu, and Ran Cui. 2019 · 2019
Later among the works it cites.
Gender Bias in Contextualized Word Embeddings
Zhao, Jieyu, Tianlu Wang, Mark Yatskar, Ryan Cotterell, Vicente Ordonez, and Kai-Wei Chang. 2019 · 2019
Later among the works it cites.
Load What You Need: Smaller Versions of Mutililingual BERT
Abdaoui, Amine, Camille Pradel, and Grégoire Sigel. 2020 · 2020
Later among the works it cites.
German’s Next Language Model
Chan, Branden, Stefan Schweter, and Timo Möller. 2020 · 2020
Later among the works it cites.
Estimation of socioeconomic attributes from location information
Doi, Shohei, Takayuki Mizuno, and Naoya Fujiwara. 2020 · 2020
Later among the works it cites.
Long Distance Relationships Without Time Travel: Boosting the Performance of a Sparse Predictive Autoencoder in Sequence Modeling
Gordon, Jeremy, David Rawlinson, and Subutai Ahmad. 2020 · 2020
Later among the works it cites.
Privacy Enhanced Multimodal Neural Representations for Emotion Recognition
Jaiswal, Mimansa and Emily Mower Provost. 2020 · 2020
Later among the works it cites.
Predicting political sentiments of voters from Twitter in multi-party contexts
Khatua, Aparup, Apalak Khatua, and Erik Cambria. 2020 · 2020
Later among the works it cites.
FlauBERT: Unsupervised Language Model Pre-training for French
Le, Hang, Loïc Vial, Jibril Frej, Vincent Segonne, Maximin Coavoux, Benjamin Lecouteux, Alexandre Allauzen, Benoit Crabbé, Laurent Besacier, and Didier Schwab. 2020 · 2020
Later among the works it cites.
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension
Lewis, Mike, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Speaker-Invariant Affective Representation Learning via Adversarial Training
Li, Haoqi, Ming Tu, Jing Huang, Shrikanth Narayanan, and Panayiotis Georgiou. 2020a · 2020
Later among the works it cites.
CamemBERT: a Tasty French Language Model
Martin, Louis, Benjamin Muller, Pedro Javier Ortiz Suárez, Yoann Dupont, Laurent Romary, Éric de la Clergerie, Djamé Seddah, and Benoît Sagot. 2020 · 2020
Later among the works it cites.
Regulatory responses to medical machine learning
Minssen, Timo, Sara Gerke, Mateo Aboy, Nicholson Price, and Glenn Cohen. 2020 · 2020
Later among the works it cites.
Participatory Research for Low-resourced Machine Translation: A Case Study in African Languages
Nekoto, Wilhelmina, Vukosi Marivate, Tshinondiwa Matsila, Timi Fasubaa, Tajudeen Kolawole, Taiwo Fagbohungbe, Solomon Oluwole Akinola, Shamsuddeen Hassan Muhammad, Salomon Kabongo, Salomey Osei, Sackey Freshia, Rubungo Andre Niyongabo, Ricky Macharm, Perez Ogayo, Orevaoghene Ahia, Musie Meressa, Mofe Adeyemi, Masabata Mokgesi-Selinga, Lawrence Okegbemi, Laura Jane Martinus, Kolawole Tajudeen, Kevin Degila, Kelechi Ogueji, Kathleen Siminyu, Julia Kreutzer, Jason Webster, Jamiil Toure Ali, Jade Abbott, Iroro Orife, Ignatius Ezeani, Idris Abdulkabir Dangana, Herman Kamper, Hady Elsahar, Goodness Duru, Ghollah Kioko, Espoir Murhabazi, Elan van Biljon, Daniel Whitenack, Christopher Onyefuluchi, Chris Emezue, Bonaventure Dossou, Blessing Sibanda, Blessing Itoro Bassey, Ayodele Olabiyi, Arshath Ramkilowan, Alp Öktem, Adewale Akinfaderin, and Abdallah Bashir. 2020 · 2020
Later among the works it cites.
A Monolingual Approach to Contextualized Word Embeddings for Mid-Resource Languages
Ortiz Suárez, Pedro Javier, Laurent Romary, and Benoît Sagot. 2020 · 2020
Later among the works it cites.
The Impact of Machine Learning on the Future of Insurance Industry
Paruchuri, Harish. 2020 · 2020
Later among the works it cites.
Beyond concern: socio-demographic and attitudinal influences on privacy and disclosure choices
Pomfret, Liam, Josephine Previte, and Len Coote. 2020 · 2020
Later among the works it cites.
Pre-trained models for natural language processing: A survey
Qiu, XiPeng, TianXiang Sun, YiGe Xu, YunFan Shao, Ning Dai, and XuanJing Huang. 2020 · 2020
Later among the works it cites.
Making Monolingual Sentence Embeddings Multilingual using Knowledge Distillation
Reimers, Nils and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
Multilingual Universal Sentence Encoder for Semantic Retrieval
Yang, Yinfei, Daniel Cer, Amin Ahmad, Mandy Guo, Jax Law, Noah Constant, Gustavo Hernandez Abrego, Steve Yuan, Chris Tar, Yun-hsuan Sung, Brian Strope, and Ray Kurzweil. 2020 · 2020
Later among the works it cites.
LocMIA: Membership Inference Attacks against Aggregated Location Data
Zhang, Guanglin, Anqi Zhang, and Ping Zhao. 2020 · 2020
Later among the works it cites.
Recent Advances in Transfer Learning for Cross-Dataset Visual Recognition: A Problem-Oriented Perspective
Zhang, Jing, Wanqing Li, Philip Ogunbona, and Dong Xu. 2020 · 2020
Later among the works it cites.
Privacy Preserving Text Representation Learning Using BERT
Alnasser, Walaa, Ghazaleh Beigi, and Huan Liu. 2021 · 2021
Later among the works it cites.
Supporting Privacy, Trust, and Personalization in Online Learning
Anwar, Mohd. 2021 · 2021
Later among the works it cites.
Quantifying Reproducibility in NLP and ML
Belz, Anya. 2021 · 2021
Later among the works it cites.
TEM: High Utility Metric Differential Privacy on Text
Carvalho, Ricardo Silva, Theodore Vasiloudis, and Oluwaseyi Feyisetan. 2021 · 2021
Later among the works it cites.
Median age - The World Factbook
Central Intelligence Agency. 2021 · 2021
Later among the works it cites.
Adversarial Stylometry in the Wild: Transferable Lexical Substitution Attacks on Author Profiling
Emmery, Chris, Ákos Kádár, and Grzegorz Chrupała. 2021 · 2021
Later among the works it cites.
Pre-Trained Models: Past, Present and Future
Han, Xu, Zhengyan Zhang, Ning Ding, Yuxian Gu, Xiao Liu, Yuqi Huo, Jiezhong Qiu, Liang Zhang, Wentao Han, Minlie Huang, Qin Jin, Yanyan Lan, Yang Liu, Zhiyuan Liu, Zhiwu Lu, Xipeng Qiu, Ruihua Song, Jie Tang, Ji-Rong Wen, Jinhui Yuan, Wayne Xin Zhao, and Jun Zhu. 2021 · 2021
Later among the works it cites.
Public attitudes towards algorithmic personalization and use of personal data online: evidence from Germany, Great Britain, and the United States
Kozyreva, Anastasia, Philipp Lorenz-Spreen, Ralph Hertwig, Stephan Lewandowsky, and Stefan M. Herzog. 2021 · 2021
Later among the works it cites.
Grammatical Gender: Acquisition, Attrition, and Change
Lohndal, Terje and Marit Westergaard. 2021 · 2021
Later among the works it cites.
Scaling Language Model Training to a Trillion Parameters Using Megatron
Narayanan, Deepak, Mohammad Shoeybi, Jared Casper, Patrick LeGresley, Mostofa Patwary, Vijay Korthikanti, Dmitri Vainbrand, and Bryan Catanzaro. 2021 · 2021
Later among the works it cites.
Data and its (dis)contents: A survey of dataset development and use in machine learning research
Paullada, Amandalynne, Inioluwa Deborah Raji, Emily M. Bender, Emily Denton, and Alex Hanna. 2021 · 2021
Later among the works it cites.
CAPE: Context-Aware Private Embeddings for Private Language Learning
Plant, Richard, Dimitra Gkatzia, and Valerio Giuffrida. 2021 · 2021
Later among the works it cites.
Case Study: Deontological Ethics in NLP
Prabhumoye, Shrimai, Brendon Boldt, Ruslan Salakhutdinov, and Alan W Black. 2021 · 2021
Later among the works it cites.
Are the Multilingual Models Better? Improving Czech Sentiment with Transformers
Přibáň, Pavel and Josef Steinberger. 2021 · 2021
Later among the works it cites.
How Good is Your Tokenizer? On the Monolingual Performance of Multilingual Language Models
Rust, Phillip, Jonas Pfeiffer, Ivan Vulić, Sebastian Ruder, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
Do not neglect related languages: The case of low-resource Occitan cross-lingual word embeddings
Woller, Lisa, Viktor Hangya, and Alexander Fraser. 2021 · 2021
Later among the works it cites.
Domain-adversarial training of neural networks
Ganin, Yaroslav, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky. 2016 · 2030
Closest in time.