2022

Are Large Pre-Trained Language Models Leaking Your Personal Information?

Huang, Jie, Shao, Hanyin, Chang, Kevin Chen-Chuan

Understand

Are Large Pre-Trained Language Models Leaking Your Personal Information? In this paper, we analyze whether Pre-Trained Language Models (PLMs) are prone to leaking personal information.

  • Specifically, we query PLMs for email addresses with contexts of the email address or prompts containing the owner's name.
  • We find that PLMs do leak personal information due to memorization.
  • However, since the models are weak at association, the risk of specific personal information being extracted by attackers is low.

Reading the bibliography…