Fetching the paper…
Reading the bibliography…
The scalability of large language models (LLMs) in handling high-complexity models and large-scale datasets has led to tremendous successes in pivotal domains.
K. Simonyan and A. Zisserman, “Very Deep Convolutional Networks for Large-scale Image Recognition,” in Proc. ICLR , 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep Residual Learning for Image Recognition,” in Proc. CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
J. Novikova, O. Dušek, and V. Rieser, “The E2E Dataset: New Challenges For End-to-End Generation,” in Proc. SIGDIAL , 2017
2017
Earlier work this paper cites.
B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas, “Communication-efficient Learning of Deep Networks From Decentralized Data,” in Proc. AISTATS , 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language Models are Unsupervised Multitask Learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-Efficient Transfer Learning for NLP,” in Proc. ICML , 2019, pp. 2790–2799
2019
Earlier work this paper cites.
M. Nasr, R. Shokri, and A. Houmansadr, “Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks Against Centralized and Federated Learning,” in Proc. SP , 2019
2019
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
R. Dale, “GPT-3: What’s It Good For?” Nat. Lang. Eng. , vol. 27, no. 1, pp. 113–118, 2021
2021
Earlier work this paper cites.
Z. Nan, H. Guan, X. Shen, and C. Liao, “Deep NLP-Based Co-evolvement for Synthesizing Code Analysis from Natural Language,” in Proc. of CC , 2021
2021
Earlier work this paper cites.
J. Pfeiffer, A. Kamath, A. Rücklé, K. Cho, and I. Gurevych, “AdapterFusion: Non-Destructive Task Composition for Transfer Learning,” in Proc. EACL , 2021, pp. 487–503
2021
Earlier work this paper cites.
R. Karimi Mahabadi, J. Henderson, and S. Ruder, “Compacter: Efficient Low-Rank Hypercomplex Adapter Layers,” 2021, pp. 1022–1035
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
A. Aghajanyan, S. Gupta, and L. Zettlemoyer, “Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning,” in Proc. IJCNLP , 2021, pp. 7319–7328
2021
Earlier work this paper cites.
D. Pasquini, G. Ateniese, and M. Bernaschi, “Unleashing the Tiger: Inference Attacks on Split Learning,” in Proc. CCS , 2021
2021
Earlier work this paper cites.
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, U. Erlingsson et al. , “Extracting Training Data From Large Language Models,” in USENIX Security , 2021, pp. 2633–2650
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
C. Thapa, P. C. M. Arachchige, S. Camtepe, and L. Sun, “Splitfed: When Federated Learning Meets Split Learning,” in Proc. AAAI , 2022
2022
Earlier work this paper cites.
Z. Lin, L. Wang, J. Ding, B. Tan, and S. Jin, “Channel power gain estimation for terahertz vehicle-to-infrastructure networks,” IEEE Commun. Lett. , vol. 27, no. 1, pp. 155–159, 2022
2022
Earlier work this paper cites.
X. Liu, Y. Deng, and T. Mahmoodi, “Wireless Distributed Learning: A New Hybrid Split and Federated Learning Approach,” IEEE Trans. Wireless Commun. , vol. 22, no. 4, pp. 2650–2665, 2022
2022
Cited alongside, same era.
Huawei, NET4AI: Supporting AI as a Service in 6G . Cambridge, U.K.: Cambridge Univ. Press, 2022
2022
Cited alongside, same era.
S. Gupta, Y. Huang, Z. Zhong, T. Gao, K. Li, and D. Chen, “Recovering Private Text in Federated Learning of Language Models,” Proc. NIPS , 2022
2022
Cited alongside, same era.
OpenAI, “GPT-4 Technical Report,” arXiv preprint arXiv:2303.08774 , 2023
2023
Cited alongside, same era.
Z. Cheng, X. Xia, M. Liwang, X. Fan, Y. Sun, X. Wang, and L. Huang, “CHEESE: Distributed Clustering-based Hybrid Federated Split Learning over Edge Networks,” IEEE Trans. Parallel Distrib. Syst. , 2023
2023
Later among the works it cites.
L. Li, Y. Zhang, and L. Chen, “Prompt Distillation for Efficient LLM-based Recommendation,” in Proc. CIKM , 2023, pp. 1348–1357
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Tian, X. Li, H. Zhang, C. Zhao, B. Li, X. Wang, and F.-Y. Wang, “VistaGPT: Generative Parallel Transformers for Vehicles with Intelligent Systems for Transport Automation,” IEEE Trans. Intell. Veh. , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
A. J. Thirunavukarasu, D. S. J. Ting, K. Elangovan, L. Gutierrez, T. F. Tan, and D. S. W. Ting, “Large Language Models in Medicine,” Nature medicine , vol. 29, no. 8, pp. 1930–1940, 2023
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
T. Ni, G. Lan, J. Wang, Q. Zhao, and W. Xu, “Eavesdropping Mobile App Activity via Radio-Frequency Energy Harvesting,” in Proc. USENIX Security Symposium , 2023
2023
Later among the works it cites.
T. Ni, X. Zhang, and Q. Zhao, “Recovering Fingerprints from In-Display Fingerprint Sensors via Electromagnetic Side Channel,” in Proc. ACM CCS , 2023
2023
Later among the works it cites.
L. Cardenas, K. Parajes, M. Zhu, and S. Zhai, “AutoHealth: Advanced LLM-Empowered Wearable Personalized Medical Butler for Parkinson’s Disease Management,” in Proc. CCWC , 2024
2024
Closest in time.
2024
Closest in time.
W. Wang, Z. Chen, X. Chen, J. Wu, X. Zhu, G. Zeng, P. Luo, T. Lu, J. Zhou, Y. Qiao et al. , “Visionllm: Large Language Model is Also An Open-ended Decoder For Vision-centric Tasks,” Proc. NIPS , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Z. Lin, G. Zhu, Y. Deng, X. Chen, Y. Gao, K. Huang, and Y. Fang, “Efficient Parallel Split Learning over Resource-constrained Wireless Edge Networks,” IEEE Trans. Mobile Comput. , 2024
2024
Closest in time.
Z. Lin, G. Qu, X. Chen, and K. Huang, “Split Learning in 6G Edge Networks,” IEEE Wireless Commun. , 2024
2024
Closest in time.
C. Cui, Y. Ma, X. Cao, W. Ye, Y. Zhou, K. Liang, J. Chen, J. Lu, Z. Yang, K.-D. Liao et al. , “A Survey on Multimodal Large Language Models for Autonomous Driving,” in Proc. WCAC , 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
G. Qu, Z. Lin, F. Liu, X. Chen, and K. Huang, “TrimCaching: Parameter-sharing AI Model Caching in Wireless Edge Networks,” in Proc. ICDCS , 2024
2024
Closest in time.
T. Dettmers, A. Pagnoni, A. Holtzman, and L. Zettlemoyer, “Qlora: Efficient Finetuning of Quantized LLMs,” Proc. NIPS , 2024
2024
Closest in time.