Fetching the paper…
Reading the bibliography…
Advanced AI assistants combine frontier LLMs and tool access to autonomously perform complex tasks on behalf of users.
“Privacy and contextual integrity: framework and applications”
A Barth, A Datta, J Mitchell and H Nissenbaum · 2006
Earlier work this paper cites.
“Privacy in context: Technology, policy, and the integrity of social life”
Helen Nissenbaum · 2009
Earlier work this paper cites.
“Privacy, Context, and Oversharing: Reputational Challenges in a Web 2.0 World”
Michael Zimmer and Anthony Hoffman · 2012
Earlier work this paper cites.
“Publication 1544 (09/2014), Reporting Cash Payments of Over $10,000”, https://www.irs.gov/publications/p1544
U.S. Internal Service · 2014
Earlier work this paper cites.
URL: https://credit.niso.org/
“CRediT – Contributor Roles Taxonomy”, 2015 · 2015
Earlier work this paper cites.
“Measuring Privacy: An Empirical Test Using Context To Expose Confounding Variables”, 2015
Kirsten Martin and Helen Nissenbaum · 2015
Earlier work this paper cites.
“Learning Privacy Expectations by Crowdsourcing Contextual Informational Norms”
Yan Shvartzshnaider, Schrasing Tong, Thomas Wies, Paula Kift, Helen Nissenbaum, Lakshminarayanan Subramanian and Prateek Mittal · 2016
Earlier work this paper cites.
“Contextual integrity up and down the data food chain”
Helen Nissenbaum · 2019
Earlier work this paper cites.
“VACCINE: Using Contextual Integrity For Data Leakage Detection”
Yan Shvartzshnaider, Zvonimir Pavlinovic, Ananth Balashankar, Thomas Wies, Lakshminarayanan Subramanian, Helen Nissenbaum and Prateek Mittal · 2019
Earlier work this paper cites.
“Privacy Norms for Smart Home Personal Assistants”
Noura Abdi, Xiao Zhan, Kopo. Ramokapane and Jose. Such · 2021
Earlier work this paper cites.
“Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences”
Denis Emelin, Ronan Le, Jena Hwang, Maxwell Forbes and Yejin Choi · 2021
Earlier work this paper cites.
“Generating Datasets with Pretrained Language Models”
Timo Schick and Hinrich Sch\"utze · 2021
Earlier work this paper cites.
“What does it mean for a language model to preserve privacy?”
Hannah Brown, Katherine Lee, Fatemehsadat Mireshghallah, Reza Shokri and Florian Tram\‘er · 2022
Earlier work this paper cites.
“Contextualizing language models for norms diverging from social majority”
Niklas Kiehne, Hermann Kroll and Wolf Balke · 2022
Earlier work this paper cites.
“Large language models are zero-shot reasoners”
Takeshi Kojima, Shixiang Gu, Machel Reid, Yutaka Matsuo and Yusuke Iwasawa · 2022
Earlier work this paper cites.
“Internet-Augmented Dialogue Generation”
Mojtaba Komeili, Kurt Shuster and Jason Weston · 2022
Cited alongside, same era.
“Runtime Permissions for Privacy in Proactive Intelligent Assistants”
Nathan Malkin, David. Wagner and Serge Egelman · 2022
Cited alongside, same era.
“TALM: Tool Augmented Language Models”
Aaron Parisi, Yao Zhao and Noah Fiedel · 2022
Cited alongside, same era.
“Moral foundations of large language models”
Marwa Abdulhai, Gregory Serapio-Garcia, Cl\’ement Crepy, Daria Valter, John Canny and Natasha Jaques · 2023
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman and Shyamal Anadkat · 2023
Cited alongside, same era.
“LLaMA: Open and Efficient Foundation Language Models”, 2023
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave and Guillaume Lample · 2023
Later among the works it cites.
“Universal and transferable adversarial attacks on aligned language models”
Andy Zou, Zifan Wang, J Kolter and Matt Fredrikson · 2023
Later among the works it cites.
“Air Gap: Protecting Privacy-Conscious Conversational Agents”
Eugene Bagdasaryan, Ren Yi, Sahra Ghalebikesabi, Peter Kairouz, Marco Gruteser, Sewoong Oh, Borja Balle and Daniel Ramage · 2024
Closest in time.
“Whispers in the Machine: Confidentiality in LLM-integrated Systems”
Jonathan Evertz, Merlin Chlosta, Lea Sch\"onherr and Thorsten Eisenhofer · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Probing Pre-Trained Language Models for Cross-Cultural Differences in Values”
Arnav Arora, Lucie-Aim\’ee Kaffee and Isabelle Augenstein · 2023
Cited alongside, same era.
“Quantifying Memorization Across Neural Language Models”
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, Katherine Lee, Florian Tram\‘er and Chiyuan Zhang · 2023
Cited alongside, same era.
“PAL: Program-aided Language Models”
Luyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon, Pengfei Liu, Yiming Yang, Jamie Callan and Graham Neubig · 2023
Cited alongside, same era.
Albert Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford and Devendra Chaplot · 2023
Cited alongside, same era.
“Scalable extraction of training data from (production) language models”
Milad Nasr, Nicholas Carlini, Jonathan Hayase, Matthew Jagielski, A Cooper, Daphne Ippolito, Christopher Choquette-Choo, Eric Wallace, Florian Tram\‘er and Katherine Lee · 2023
Cited alongside, same era.
“ChatGPT plugins - OpenAI”, 2023
OpenAI · 2023
Cited alongside, same era.
“Identifying the risks of lm agents with an lm-emulated sandbox”
Yangjun Ruan, Honghua Dong, Andrew Wang, Silviu Pitis, Yongchao Zhou, Jimmy Ba, Yann Dubois, Chris Maddison and Tatsunori Hashimoto · 2023
Cited alongside, same era.
Iason Gabriel, Arianna Manzini, Geoff Keeling, Lisa Hendricks, Verena Rieser, Hasan Iqbal, Nenad Tomašev and Ira Ktena · 2024
Closest in time.
“LLM Censorship: The Problem and its Limitations”
David Glukhov, Ilia Shumailov, Yarin Gal, Nicolas Papernot and Vardan Papyan · 2024
Closest in time.
“Gemini for Google Workspace”, 2024
Google · 2024
Closest in time.
“Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory”
Niloofar Mireshghallah, Hyunwoo Kim, Xuhui Zhou, Yulia Tsvetkov, Maarten Sap, Reza Shokri and Yejin Choi · 2024
Closest in time.
“Evaluating the moral beliefs encoded in llms”
Nino Scherrer, Claudia Shi, Amir Feder and David Blei · 2024
Closest in time.
“Toolformer: Language models can teach themselves to use tools”
Timo Schick, Jane Dwivedi-Yu, Roberto Dessi, Roberta Raileanu, Maria Lomeli, Eric Hambro, Luke Zettlemoyer, Nicola Cancedda and Thomas Scialom · 2024
Closest in time.
“Gemma 2: Improving Open Language Models at a Practical Size”
Gemma Team, Morgane Riviere, Shreya Pathak, Pier Sessa, Cassidy Hardin, Surya Bhupatiraju, L\’eonard Hussenot, Thomas Mesnard, Bobak Shahriari and Alexandre Ram\’e · 2024
Closest in time.
“Privacy”
Andrew Trask, Geoff Keeling, Borja Balle, Sarah de Haas, Yetunde Ibitoye and Iason Gabriel · 2024
Closest in time.
“The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions”
Eric Wallace, Kai Xiao, Reimar Leike, Lilian Weng, Johannes Heidecke and Alex Beutel · 2024
Closest in time.
“Measuring Social Norms of Large Language Models”
Ye Yuan, Kexin Tang, Jianhao Shen, Ming Zhang and Chenguang Wang · 2024
Closest in time.