Fetching the paper…
Reading the bibliography…
Security of model parameters and user data is critical for Transformer-based services, such as ChatGPT.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 1910
Earlier work this paper cites.
Privacy-preserving clustering by object similarity-based representation and dimensionality reduction transformation. In Proceedings of the 2004 ICDM Workshop on Privacy and Security Aspects of Data Mining . 40–46
Stanley RM Oliveira and Osmar R Zaiane. 2004 · 2004
Earlier work this paper cites.
Measuring and testing dependence by correlation of distances
Gábor J Székely, Maria L Rizzo, and Nail K Bakirov. 2007 · 2007
Earlier work this paper cites.
Cryptographic Complexity of Multi-Party Computation Problems: Classifications and Separations. In Advances in Cryptology – CRYPTO 2008 , David Wagner (Ed.). Springer Berlin Heidelberg, Berlin, Heidelberg, 262–279
Manoj Prabhakaran and Mike Rosulek. 2008 · 2008
Earlier work this paper cites.
Analysis of one-time random projections for privacy preserving compressed sensing
Tiziano Bianchi, Valerio Bioglio, and Enrico Magli. 2015 · 2015
Earlier work this paper cites.
A novel semi-symmetric encryption algorithm for internet applications
N Fares and Shavan Askar. 2016 · 2016
Earlier work this paper cites.
Information verification cryptosystem using one-time keys based on double random phase encoding and public-key cryptography
Tieyu Zhao, Qiwen Ran, Lin Yuan, Yingying Chi, and Jing Ma. 2016 · 2016
Earlier work this paper cites.
A new algorithm combining substitution & transposition cipher techniques for secure communication. In 2017 International Conference on Trends in Electronics and Informatics (ICEI) . IEEE, 619–624
Umang Bhargava, Aparna Sharma, Raghav Chawla, and Prateek Thakral. 2017 · 2017
Earlier work this paper cites.
A semi-symmetric image encryption scheme based on the function projective synchronization of two hyperchaotic systems
Xiaoqiang Di, Jinqing Li, Hui Qi, Ligang Cong, and Huamin Yang. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
Property inference attacks on fully connected neural networks using permutation invariant representations. In Proceedings of the 2018 ACM SIGSAC conference on computer and communications security . 619–633
Karan Ganju, Qi Wang, Wei Yang, Carl A Gunter, and Nikita Borisov. 2018 · 2018
Earlier work this paper cites.
Privacy-preserving svm computing in the encrypted domain. In 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) . IEEE, 897–902
Takahiro Maekawa, Ayana Kawamura, Yuma Kinoshita, and Hitoshi Kiya. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al · 2018
Earlier work this paper cites.
The annotated transformer. In Proceedings of workshop for NLP open source software (NLP-OSS) . 52–60
Alexander M Rush. 2018 · 2018
Earlier work this paper cites.
A theoretical analysis of noisy sparse subspace clustering on dimensionality-reduced data
Yining Wang, Yu-Xiang Wang, and Aarti Singh. 2018 · 2018
Earlier work this paper cites.
Generating long sequences with sparse transformers
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever. 2019 · 2019
Earlier work this paper cites.
Nemo: a toolkit for building ai applications using neural modules
Oleksii Kuchaiev, Jason Li, Huyen Nguyen, Oleksii Hrinchuk, Ryan Leary, Boris Ginsburg, Samuel Kriman, Stanislav Beliaev, Vitaly Lavrukhin, Jack Cook, et al · 2019
Earlier work this paper cites.
Set transformer: A framework for attention-based permutation-invariant neural networks. In International conference on machine learning . PMLR, 3744–3753
Juho Lee, Yoonho Lee, Jungtaek Kim, Adam Kosiorek, Seungjin Choi, and Yee Whye Teh. 2019 · 2019
Earlier work this paper cites.
Language Models are Unsupervised Multitask Learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019a · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Root mean square layer normalization
Biao Zhang and Rico Sennrich. 2019 · 2019
Cited alongside, same era.
Can we use split learning on 1d cnn models for privacy preserving training?. In Proceedings of the 15th ACM Asia Conference on Computer and Communications Security . 305–318
Sharif Abuadbba, Kyuyeon Kim, Minki Kim, Chandra Thapa, Seyit A Camtepe, Yansong Gao, Hyoungshick Kim, and Surya Nepal. 2020 · 2020
Cited alongside, same era.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan. 2020 · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Splitfed: When federated learning meets split learning. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 36. 8485–8493
Chandra Thapa, Pathum Chamikara Mahawaga Arachchige, Seyit Camtepe, and Lichao Sun. 2022 · 2022
Later among the works it cites.
Protect privacy from gradient leakage attack in federated learning. In IEEE INFOCOM 2022-IEEE Conference on Computer Communications . IEEE, 580–589
Junxiao Wang, Song Guo, Xin Xie, and Heng Qi. 2022 · 2022
Later among the works it cites.
Towards secure and practical machine learning via secret sharing and random permutation
Fei Zheng, Chaochao Chen, Xiaolin Zheng, and Mingjie Zhu. 2022 · 2022
Later among the works it cites.
Microsoft: We’re bringing ChatGPT to the Azure cloud-computing service
2023 · 2023
Closest in time.
ChatGPT is now available in Azure OpenAI Service
Azure. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Transnet: Training privacy-preserving neural network over transformed layer
Qijian He, Wei Yang, Bingren Chen, Yangyang Geng, and Liusheng Huang. 2020 · 2020
Cited alongside, same era.
Privacy risks of general-purpose language models. In 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 1314–1331
Xudong Pan, Mi Zhang, Shouling Ji, and Min Yang. 2020 · 2020
Cited alongside, same era.
Glu variants improve transformer
Noam Shazeer. 2020 · 2020
Cited alongside, same era.
Deep compressive offloading: Speeding up neural network inference by trading edge computation for network latency. In Proceedings of the 18th conference on embedded networked sensor systems . 476–488
Shuochao Yao, Jinyang Li, Dongxin Liu, Tianshi Wang, Shengzhong Liu, Huajie Shao, and Tarek Abdelzaher. 2020 · 2020
Cited alongside, same era.
Coedge: Cooperative dnn inference with adaptive workload partitioning over heterogeneous edge devices
Liekang Zeng, Xu Chen, Zhi Zhou, Lei Yang, and Junshan Zhang. 2020 · 2020
Cited alongside, same era.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby. 2021 · 2021
Cited alongside, same era.
LoRA: Low-Rank Adaptation of Large Language Models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Samuel Carreira, Tomás Marques, José Ribeiro, and Carlos Grilo. 2023 · 2023
Closest in time.
On-Device Deep Learning for Mobile and Wearable Sensing Applications: A Review
Ozlem Durmaz Incel and Sevda Özge Bursa. 2023 · 2023
Closest in time.
Samsung Bans ChatGPT Among Employees After Sensitive Code Leak
Forbes. 2023 · 2023
Closest in time.
Ciphergpt: Secure two-party gpt inference
Xiaoyang Hou, Jian Liu, Jingyu Li, Yuhan Li, Wen-jie Lu, Cheng Hong, and Kui Ren. 2023 · 2023
Closest in time.
Open-LLM-Leaderboard
HuggingFace. 2023 · 2023
Closest in time.
High-Efficiency Device-Cloud Collaborative Transformer Model. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2203–2209
Penghao Jiang, Ke Xin, Chunxi Li, and Yinsi Zhou. 2023 · 2023
Closest in time.
Visual Instruction Tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2023 · 2023
Closest in time.
Binarizing split learning for data privacy enhancement and computation reduction
Ngoc Duy Pham, Alsharif Abuadbba, Yansong Gao, Tran Khoa Phan, and Naveen Chilamkurti. 2023 · 2023
Closest in time.
ChatGPT sets record for fastest-growing user base - analyst note
REUTERS. 2023 · 2023
Closest in time.
TPTU: Large Language Model-based AI Agents for Task Planning and Tool Usage
Jingqing Ruan, Yihong Chen, Bin Zhang, Zhiwei Xu, Tianpeng Bao, Guoqing Du, Shiwei Shi, Hangyu Mao, Ziyue Li, Xingyu Zeng, and Rui Zhao. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Closest in time.
Smoothquant: Accurate and efficient post-training quantization for large language models. In International Conference on Machine Learning . PMLR, 38087–38099
Guangxuan Xiao, Ji Lin, Mickael Seznec, Hao Wu, Julien Demouth, and Song Han. 2023 · 2023
Closest in time.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, et al · 2023
Closest in time.
Albert Q. Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, Gianna Lengyel, Guillaume Bour, Guillaume Lample, Lélio Renard Lavaud, Lucile Saulnier, Marie-Anne Lachaux, Pierre Stock, Sandeep Subramanian, Sophia Yang, Szymon Antoniak, Teven Le Scao, Théophile Gervet, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed. 2024 · 2024
Closest in time.
tiktoken
OpenAI. 2024 · 2024
Closest in time.