Fetching the paper…
Reading the bibliography…
Self-disclosure, while being common and rewarding in social media interaction, also poses privacy risks.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al. 2019 · 1910
Earlier work this paper cites.
The hungarian method for the assignment problem
Harold W Kuhn. 1955 · 1955
Earlier work this paper cites.
Self-disclosure: An experimental analysis of the transparent self
Sidney M Jourard. 1971 · 1971
Earlier work this paper cites.
Self-disclosure: a literature review
Paul C. Cozby. 1973 · 1973
Earlier work this paper cites.
Self-disclosure topic model for classifying and analyzing twitter conversations
JinYeong Bak, Chin-Yew Lin, and Alice Oh. 2014 · 1996
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Privacy in crisis: A study of self-disclosure during the coronavirus pandemic
Taylor Blose, Prasanna Umar, Anna Squicciarini, and Sarah Rajtmajer. 2020 · 2004
Earlier work this paper cites.
Reliability in content analysis: Some common misconceptions and recommendations
Klaus Krippendorff. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Loose tweets: an analysis of privacy leaks on twitter
Huina Mao, Xin Shuai, and Apu Kapadia. 2011 · 2011
Earlier work this paper cites.
Privacy detective: Detecting private information and collective privacy behavior in a large social network
Aylin Caliskan Islam, Jonathan Walsh, and Rachel Greenstadt. 2014 · 2014
Earlier work this paper cites.
Detecting and characterizing mental health related self-disclosure in social media
Sairam Balani and Munmun De Choudhury. 2015 · 2015
Earlier work this paper cites.
An analysis of the user occupational class through twitter content
Daniel Preoţiuc-Pietro, Vasileios Lampos, and Nikolaos Aletras. 2015 · 2015
Earlier work this paper cites.
An analysis of domestic abuse discourse on reddit
Nicolas Schrading, Cecilia Ovesdotter Alm, Raymond Ptucha, and Christopher Homan. 2015 · 2015
Earlier work this paper cites.
Understanding social media disclosures of sexual abuse through the lenses of support seeking and anonymity
Nazanin Andalibi, Oliver L Haimson, Munmun De Choudhury, and Andrea Forte. 2016 · 2016
Earlier work this paper cites.
Fasttext.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hérve Jégou, and Tomas Mikolov. 2016 · 2016
Earlier work this paper cites.
Regulation (eu) 2016/679 of the european parliament and of the council
Protection Regulation. 2016 · 2016
Earlier work this paper cites.
Everyday online sharing
Manya Sleeper. 2016 · 2016
Earlier work this paper cites.
Modeling self-disclosure in social networking sites
Yi-Chia Wang, Moira Burke, and Robert Kraut. 2016 · 2016
Earlier work this paper cites.
Multitask learning for mental health conditions with limited social media data
Adrian Benton, Margaret Mitchell, Dirk Hovy, et al. 2017 · 2017
Earlier work this paper cites.
De-identification of patient notes with recurrent neural networks
Franck Dernoncourt, Ji Young Lee, Ozlem Uzuner, and Peter Szolovits. 2017 · 2017
Earlier work this paper cites.
Detecting personal medication intake in twitter: an annotated corpus and baseline classification system
Ari Klein, Abeed Sarker, Masoud Rouhizadeh, Karen O’Connor, and Graciela Gonzalez. 2017 · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2017 · 2017
Earlier work this paper cites.
Self-disclosure and channel difference in online health support groups
Diyi Yang, Zheng Yao, and Robert Kraut. 2017 · 2017
Earlier work this paper cites.
Depression and self-harm risk assessment in online forums
Andrew Yates, Arman Cohan, and Nazli Goharian. 2017 · 2017
Earlier work this paper cites.
Content analysis: An introduction to its methodology
Klaus Krippendorff. 2018 · 2018
Earlier work this paper cites.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Earlier work this paper cites.
AnonyMate: A toolkit for anonymizing unstructured chat data
Allison Adams, Eric Aili, Daniel Aioanei, Rebecca Jonsson, Lina Mickelsson, Dagmar Mikmekova, Fred Roberts, Javier Fernandez Valencia, and Roger Wechsler. 2019 · 2019
Earlier work this paper cites.
Speak up, fight back! detection of social media disclosures of sexual harassment
Arijit Ghosh Chowdhury, Ramit Sawhney, Puneet Mathur, Debanjan Mahata, and Rajiv Ratn Shah. 2019 · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al. 2019 · 2019
Cited alongside, same era.
Detection and analysis of self-disclosure in online news commentaries
Prasanna Umar, Anna Squicciarini, and Sarah Rajtmajer. 2019 · 2019
Cited alongside, same era.
Identifying adverse drug events mentions in tweets using attentive, collocated, and aggregated medical representation
Xinyan Zhao, Deahan Yu, and VG Vinod Vydiswaran. 2019 · 2019
Cited alongside, same era.
A semantics-based approach to disclosure classification in user-generated online content
Chandan Akiti, Anna Squicciarini, and Sarah Rajtmajer. 2020 · 2020
Cited alongside, same era.
Unsupervised text deidentification
John Morris, Justin Chiu, Ramin Zabih, and Alexander M Rush. 2022 · 2022
Later among the works it cites.
The text anonymization benchmark (tab): A dedicated corpus and evaluation framework for text anonymization
Ildikó Pilán, Pierre Lison, Lilja Øvrelid, Anthi Papadopoulou, David Sánchez, and Montserrat Batet. 2022 · 2022
Later among the works it cites.
CometKiwi: IST-unbabel 2022 submission for the quality estimation shared task
Ricardo Rei, Marcos Treviso, Nuno M. Guerreiro, Chrysoula Zerva, Ana C Farinha, Christine Maroti, José G. C. de Souza, Taisiya Glushkova, Duarte Alves, Luisa Coheur, Alon Lavie, and André F. T. Martins. 2022 · 2022
Later among the works it cites.
Measuring the language of self-disclosure across corpora
Ann-Katrin Reuel, Sebastian Peralta, João Sedoc, Garrick Sherman, and Lyle Ungar. 2022 · 2022
Later among the works it cites.
Multilingual detection of personal employment status on twitter
Manuel Tonneau, Dhaval Adjodah, Joao Palotti, Nir Grinberg, and Samuel Fraiberger. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
spaCy: Industrial-strength natural language processing in python
Matthew Honnibal, Ines Montani, Sofie Van Landeghem, Adriane Boyd, et al. 2020 · 2020
Cited alongside, same era.
Privacy and activism in the transgender community
Ada Lerner, Helen Yuxun He, Anna Kawakami, Silvia Catherine Zeamer, and Roberto Hoyle. 2020 · 2020
Cited alongside, same era.
Self-disclosure and social media: motivations, mechanisms and psychological well-being
Mufan Luo and Jeffrey T Hancock. 2020 · 2020
Cited alongside, same era.
LUKE: Deep contextualized entity representations with entity-aware self-attention
Ikuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda, and Yuji Matsumoto. 2020 · 2020
Cited alongside, same era.
PHICON: Improving generalization of clinical text de-identification models via data augmentation
Xiang Yue and Shuang Zhou. 2020 · 2020
Cited alongside, same era.
Extracting training data from large language models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al. 2021 · 2021
Cited alongside, same era.
All that’s ‘human’ is not gold: Evaluating human evaluation of generated text
Elizabeth Clark, Tal August, Sofia Serrano, Nikita Haduong, Suchin Gururangan, and Noah A. Smith. 2021 · 2021
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Later among the works it cites.
Promptner: Prompting for named entity recognition
Dhananjay Ashok and Zachary C Lipton. 2023 · 2023
Closest in time.
Amazon comprehend
AWS. 2023 · 2023
Closest in time.
Azure AI language
Azure. 2023 · 2023
Closest in time.
Can language models be instructed to protect personal information?
Yang Chen, Ethan Mendes, Sauvik Das, Wei Xu, and Alan Ritter. 2023 · 2023
Closest in time.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli. 2023 · 2023
Closest in time.
Dancing between success and failure: Edit-level simplification evaluation using SALSA
David Heineman, Yao Dou, Mounica Maddela, and Wei Xu. 2023 · 2023
Closest in time.
Training data extraction from pre-trained language models: A survey
Shotaro Ishihara. 2023 · 2023
Closest in time.
ProPILE: Probing privacy leakage in large language models
Siwon Kim, Sangdoo Yun, Hwaran Lee, Martin Gubri, Sungroh Yoon, and Seong Joon Oh. 2023 · 2023
Closest in time.
Online self-disclosure, social support, and user engagement during the covid-19 pandemic
Jooyoung Lee, Sarah Rajtmajer, Eesha Srivatsavaya, and Shomir Wilson. 2023 · 2023
Closest in time.
Analyzing leakage of personally identifiable information in language models
Nils Lukas, Ahmed Salem, Robert Sim, Shruti Tople, Lukas Wutschitz, and Santiago Zanella-Béguelin. 2023 · 2023
Closest in time.
LENS: A learnable evaluation metric for text simplification
Mounica Maddela, Yao Dou, David Heineman, and Wei Xu. 2023 · 2023
Closest in time.
Niloofar Mireshghallah, Hyunwoo Kim, Xuhui Zhou, Yulia Tsvetkov, Maarten Sap, Reza Shokri, and Yejin Choi. 2023 · 2023
Closest in time.
Crowdsourcing on sensitive data with privacy-preserving text rewriting
Nina Mouhammad, Johannes Daxenberger, Benjamin Schiller, and Ivan Habernal. 2023 · 2023
Closest in time.
How to dp-fy ml: A practical guide to machine learning with differential privacy
Natalia Ponomareva, Hussein Hazimeh, Alex Kurakin, Zheng Xu, Carson Denison, H Brendan McMahan, Sergei Vassilvitskii, Steve Chien, and Abhradeep Guha Thakurta. 2023 · 2023
Closest in time.
Identifying and mitigating privacy risks stemming from language models: A survey
Victoria Smith, Ali Shahin Shamsabadi, Carolyn Ashurst, and Adrian Weller. 2023 · 2023
Closest in time.
Beyond memorization: Violating privacy via inference with large language models
Robin Staab, Mark Vero, Mislav Balunović, and Martin Vechev. 2023 · 2023
Closest in time.
Detecting personal information in training corpora: an analysis
Nishant Subramani, Sasha Luccioni, Jesse Dodge, and Margaret Mitchell. 2023 · 2023
Closest in time.
Does fine-tuning gpt-3 with the openai api leak personally-identifiable information?
Albert Yu Sun, Eliott Zemour, Arushi Saxena, Udith Vaidyanathan, Eric Lin, Christian Lau, and Vaikkunth Mugunthan. 2023 · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Closest in time.
What clued the ai doctor in? on the influence of data source and quality for transformer-based medical self-disclosure detection
Mina Valizadeh, Xing Qian, Pardis Ranjbar-Noiey, Cornelia Caragea, and Natalie Parde. 2023 · 2023
Closest in time.
Siren’s song in the ai ocean: a survey on hallucination in large language models
Yue Zhang, Yafu Li, Leyang Cui, Deng Cai, Lemao Liu, Tingchen Fu, Xinting Huang, Enbo Zhao, Yu Zhang, Yulong Chen, et al. 2023 · 2023
Closest in time.
Are large pre-trained language models leaking your personal information?
Jie Huang, Hanyin Shao, and Kevin Chen-Chuan Chang. 2022 · 2047
Closest in time.
Discovering shifts to suicidal ideation from mental health content in social media
Munmun De Choudhury, Emre Kiciman, Mark Dredze, Glen Coppersmith, and Mrinal Kumar. 2016 · 2098
Closest in time.