Fetching the paper…
Reading the bibliography…
Optimizing the deployment of large language models (LLMs) in edge computing environments is critical for enhancing privacy and computational efficiency.
M. Nakagami, The m-Distribution—A General Formula of Intensity Distribution of Rapid Fading . Oxford: Pergamon, 1960, pp. 3–36. [Online]. Available: https://www.sciencedirect.com/science/article/pii/B9780080093062500054
1960
Earlier work this paper cites.
M. Stone, “Cross-validatory choice and assessment of statistical predictions,” Journal of the Royal Statistical Society: Series B (Methodological) , vol. 36, no. 2, pp. 111–133, 1974
1974
Earlier work this paper cites.
W. S. Cleveland, “Robust locally weighted regression and smoothing scatterplots,” Journal of the American Statistical Association , vol. 74, no. 368, pp. 829–836, 1979
1979
Earlier work this paper cites.
N. Beaulieu and C. Cheng, “Efficient Nakagami-m fading channel simulation,” IEEE Transactions on Vehicular Technology , vol. 54, no. 2, pp. 413–424, 2005
2005
Earlier work this paper cites.
M. Satyanarayanan, P. Bahl, R. Caceres, and N. Davies, “The case for vm-based cloudlets in mobile computing,” IEEE Pervasive Computing , vol. 8, no. 4, pp. 14–23, 2009
2009
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in Proceedings of the 28th International Conference on Machine Learning (ICML-11) , 2011, pp. 465–472
2011
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
S. Merity, C. Xiong, J. Bradbury, and R. Socher, “Pointer sentinel mixture models,” 2016
2016
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” Proceedings of the 33rd International Conference on Machine Learning (ICML) , vol. 48, no. 1, pp. 1928–1937, 2016. [Online]. Available: http://proceedings.mlr.press/v48/mnih16.html
2016
Earlier work this paper cites.
Y. Mao, C. You, J. Zhang, K. Huang, and K. B. Letaief, “A survey on mobile edge computing: The communication perspective,” IEEE Communications Surveys & Tutorials , vol. 19, no. 4, pp. 2322–2358, 2017
2017
Earlier work this paper cites.
P. Mach and Z. Becvar, “Mobile edge computing: A survey on architecture and computation offloading,” IEEE Communications Surveys & Tutorials , vol. 19, no. 3, pp. 1628–1656, 2017
2017
Earlier work this paper cites.
Y. Li, “Deep reinforcement learning: An overview,” 2017
2017
Earlier work this paper cites.
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Proximal policy optimization algorithms,” 2017
2017
Earlier work this paper cites.
N. Abbas, Y. Zhang, A. Taherkordi, and T. Skeie, “Mobile edge computing: A survey,” IEEE Internet of Things Journal , vol. 5, no. 1, pp. 450–465, 2018
2018
Earlier work this paper cites.
O. Gupta and R. Raskar, “Distributed learning of deep neural network over multiple agents,” Journal of Network and Computer Applications , vol. 116, no. 1, pp. 1–8, 2018
2018
Earlier work this paper cites.
J. Romoff, P. Henderson, A. Piché, V. Francois-Lavet, and J. Pineau, “Reward estimation for variance reduction in deep reinforcement learning,” 2018
2018
Earlier work this paper cites.
E. Li, L. Zeng, Z. Zhou, and X. Chen, “Edge AI: On-demand accelerating deep neural network inference via edge computing,” IEEE Transactions on Wireless Communications , vol. 19, no. 1, pp. 447–457, 2019
2019
Earlier work this paper cites.
N. C. Luong, D. T. Hoang, S. Gong, D. Niyato, P. Wang, Y.-C. Liang, and D. I. Kim, “Applications of deep reinforcement learning in communications and networking: A survey,” IEEE Communications Surveys & Tutorials , vol. 21, no. 4, pp. 3133–3174, 2019
2019
Earlier work this paper cites.
Y. Qian, J. Wu, R. Wang, F. Zhu, and W. Zhang, “Survey on reinforcement learning applications in communication networks,” Journal of Communications and Information Networks , vol. 4, no. 2, pp. 30–39, 2019
2019
Earlier work this paper cites.
L. Kaiser, M. Babaeizadeh, P. Milos, B. Osinski, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine et al. , “Model-based reinforcement learning for Atari,” 2019
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in Neural Information Processing Systems , vol. 33, no. 1, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
Q.-V. Pham, F. Fang, V. N. Ha, M. J. Piran, M. Le, L. B. Le, W.-J. Hwang, and Z. Ding, “A survey of multi-access edge computing in 5G and beyond: Fundamentals, technology integration, and state-of-the-art,” IEEE Access , vol. 8, no. 1, pp. 116 974–117 017, 2020
2020
Earlier work this paper cites.
D. Liu, C. Sun, C. Yang, and L. Hanzo, “Optimizing wireless systems using unsupervised and reinforced-unsupervised deep learning,” IEEE Network , vol. 34, no. 4, pp. 270–277, 2020
2020
Earlier work this paper cites.
J. Wei, M. Bosma, V. Y. Zhao, K. Guu, A. W. Yu, B. Lester, N. Du, A. M. Dai, and Q. V. Le, “Finetuned language models are zero-shot learners,” 2021
2021
Cited alongside, same era.
K. B. Letaief, Y. Shi, J. Lu, and J. Lu, “Edge artificial intelligence for 6G: Vision, enabling technologies, and applications,” IEEE Journal on Selected Areas in Communications , vol. 40, no. 1, pp. 5–36, 2021
2021
Cited alongside, same era.
M. Chen, D. Gündüz, K. Huang, W. Saad, M. Bennis, A. V. Feljan, and H. V. Poor, “Distributed learning in wireless networks: Recent progress and future challenges,” IEEE Journal on Selected Areas in Communications , vol. 39, no. 12, pp. 3579–3605, 2021
2021
Cited alongside, same era.
Q. Lan, Q. Zeng, P. Popovski, D. Gündüz, and K. Huang, “Progressive feature transmission for split inference at the wireless edge,” 2021
2021
Cited alongside, same era.
L. Qiao and Y. Zhou, “Timely split inference in wireless networks: An accuracy-freshness tradeoff,” IEEE Transactions on Vehicular Technology , vol. 72, no. 12, pp. 16 817–16 822, 2023
2023
Later among the works it cites.
Y. Wang, K. Guo, W. Hong, Q. Mu, and Z. Zhao, “Split learning in wireless networks: A communication and computation adaptive scheme,” in 2023 IEEE/CIC International Conference on Communications in China (ICCC) . IEEE, 2023, pp. 1–6
2023
Later among the works it cites.
T. M. Moerland, J. Broekens, A. Plaat, C. M. Jonker et al. , “Model-based reinforcement learning: A survey,” Foundations and Trends® in Machine Learning , vol. 16, no. 1, pp. 1–118, 2023
2023
Later among the works it cites.
R. T. Icarte, T. Q. Klassen, R. Valenzano, M. P. Castro, E. Waldie, and S. A. McIlraith, “Learning reward machines: A study in partially observable reinforcement learning,” Artificial Intelligence , vol. 323, no. 1, p. 103989, 2023
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
N. Shlezinger, N. Farsad, Y. C. Eldar, and A. J. Goldsmith, “Model-based machine learning for communications,” 2021
2021
Cited alongside, same era.
Y. Bai, A. Jones, K. Ndousse, A. Askell, A. Chen, N. DasSarma, D. Drain, S. Fort, D. Ganguli, T. Henighan et al. , “Training a helpful and harmless assistant with reinforcement learning from human feedback,” 2022
2022
Cited alongside, same era.
T. Le Scao, A. Fan, C. Akiki, E. Pavlick, S. Ilić, D. Hesslow, R. Castagné, A. S. Luccioni, F. Yvon, M. Gallé et al. , “Bloomloom: A 176b-pb-parameter open-access multilingual language model,” 2022
2022
Cited alongside, same era.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” 2022
2022
Cited alongside, same era.
J. Ryu, D. Won, and Y. Lee, “A study of split learning model.” in IMCOM , 2022, pp. 1–4
2022
Cited alongside, same era.
J. Karjee, P. Naik, K. Anand, and V. N. Bhargav, “Split computing: DNN inference partition with load balancing in IoT-edge platform for beyond 5G,” Measurement: Sensors , vol. 23, no. 1, p. 100409, 2022
2022
Cited alongside, same era.
L. Zhu, G. Takami, M. Kawahara, H. Kanokogi, and T. Matsubara, “Alleviating parameter-tuning burden in reinforcement learning for large-scale process control,” Computers & Chemical Engineering , vol. 158, no. 1, p. 107658, 2022
2022
Cited alongside, same era.
X. Li, L. Lu, W. Ni, A. Jamalipour, D. Zhang, and H. Du, “Federated multi-agent deep reinforcement learning for resource allocation of vehicle-to-vehicle communications,” IEEE Transactions on Vehicular Technology , vol. 71, no. 8, pp. 8810–8824, 2022
2022
Cited alongside, same era.
K. Yang, C. Shen, J. Yang, S.-p. Yeh, and J. Sydir, “Offline reinforcement learning for wireless network optimization with mixture datasets,” 2023
2023
Later among the works it cites.
C.-H. Ke and L. Astuti, “Applying multi-agent deep reinforcement learning for contention window optimization to enhance wireless network performance,” ICT Express , vol. 9, no. 5, pp. 776–782, 2023
2023
Later among the works it cites.
A. Q. Jiang, A. Sablayrolles, A. Mensch, C. Bamford, D. S. Chaplot, D. d. l. Casas, F. Bressand, G. Lengyel, G. Lample, L. Saulnier et al. , “Mistral 7B,” 2023
2023
Later among the works it cites.
G. Wang, S. Cheng, X. Zhan, X. Li, S. Song, and Y. Liu, “Openchat: Advancing open-source language models with mixed-quality data,” 2023
2023
Later among the works it cites.
X. Zhang, B. Yu, H. Yu, Y. Lv, T. Liu, F. Huang, H. Xu, and Y. Li, “Wider and deeper LLM networks are fairer LLM evaluators,” 2023
2023
Later among the works it cites.
OpenAI, “GPT-4 technical report,” OpenAI, 2024. [Online]. Available: https://cdn.openai.com/papers/gpt-4.pdf
2024
Closest in time.
M. Jin, Q. Yu, C. Zhang, D. Shu, S. Zhu, M. Du, Y. Zhang, and Y. Meng, “Health-LLM: Personalized retrieval-augmented disease prediction model,” 2024
2024
Closest in time.
Y. Chen, R. Li, Z. Zhao, C. Peng, J. Wu, E. Hossain, and H. Zhang, “NetGPT: An AI-native network architecture for provisioning beyond personalized generative services,” IEEE Network , March 2024, early Access
2024
Closest in time.
Z. Lin, G. Qu, Q. Chen, X. Chen, Z. Chen, and K. Huang, “Pushing large language models to the 6G edge: Vision, challenges, and opportunities,” 2024
2024
Closest in time.
R. Patil and V. Gudivada, “A review of current trends, techniques, and challenges in large language models (llms),” Applied Sciences , vol. 14, no. 5, p. 2074, 2024
2024
Closest in time.
Z. Lin, G. Qu, X. Chen, and K. Huang, “Split learning in 6G edge networks,” 2024
2024
Closest in time.
B. Lin, T. Peng, C. Zhang, M. Sun, L. Li, H. Zhao, W. Xiao, Q. Xu, X. Qiu, S. Li et al. , “Infinite-LLM: Efficient LLM service for long context with distattention and distributed kvcache,” 2024
2024
Closest in time.
L. Chen, N. K. Ahmed, A. Dutta, A. Bhattacharjee, S. Yu, Q. I. Mahmud, W. Abebe, H. Phan, A. Sarkar, B. Butler et al. , “Position paper: The landscape and challenges of HPC research and LLMs,” 2024
2024
Closest in time.
Q. Dong, X. Chen, and M. Satyanarayanan, “Creating edge AI from cloud-based LLMs,” in Proceedings of the 25th International Workshop on Mobile Computing Systems and Applications , 2024, pp. 8–13, early Access
2024
Closest in time.
M. Zhang, J. Cao, X. Shen, and Z. Cui, “Edgeshard: Efficient llm inference via collaborative edge computing,” 2024
2024
Closest in time.
I. Ong, “Efficient distributed LLM inference with dynamic partitioning,” California, Berkeley, Technical Report UCB/EECS-2024-108, May 2024. [Online]. Available: http://www2.eecs.berkeley.edu/Pubs/TechRpts/2024/EECS-2024-108.html
2024
Closest in time.
A. Üstün, V. Aryabumi, Z.-X. Yong, W.-Y. Ko, D. D’souza, G. Onilude, N. Bhandari, S. Singh, H.-L. Ooi, A. Kayid et al. , “Aya model: An instruction finetuned open-access multilingual language model,” 2024
2024
Closest in time.
R. Gupta and N. Sosio, “Introducing Prem-1B,” PremAI, 2024. [Online]. Available: https://blog.premai.io/introducing-prem-1b/
2024
Closest in time.