Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) exhibit potential artificial generic intelligence recently, however, their usage is costly with high response latency.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Routing to the expert: Efficient reward-guided ensemble of large language models
Keming Lu, Hongyi Yuan, Runji Lin, Junyang Lin, Zheng Yuan, Chang Zhou, and Jingren Zhou. 2024 · 1974
Earlier work this paper cites.
A contextual-bandit approach to personalized news article recommendation
Lihong Li, Wei Chu, John Langford, and Robert E Schapire. 2010 · 2010
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2020 · 2020
Earlier work this paper cites.
Multi-facet contextual bandits: A neural network perspective
Yikun Ban, Jingrui He, and Curtiss B Cook. 2021 · 2021
Earlier work this paper cites.
Ul2: Unifying language learning paradigms
Yi Tay, Mostafa Dehghani, Vinh Q Tran, Xavier Garcia, Jason Wei, Xuezhi Wang, Hyung Won Chung, Siamak Shakeri, Dara Bahri, Tal Schuster, et al. 2022 · 2022
Earlier work this paper cites.
A hierarchal bert structure for native speaker writing detection
Xinyuan Wang, Qinke Peng, Xu Mou, Haozhou Li, and Ying Wang. 2022 · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Earlier work this paper cites.
Frugalgpt: How to use large language models while reducing cost and improving performance
Lingjiao Chen, Matei Zaharia, and James Zou. 2023 · 2023
Cited alongside, same era.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al. 2023 · 2023
Cited alongside, same era.
Sehf: A summary-enhanced hierarchical framework for financial report sentiment analysis
Haozhou Li, Qinke Peng, Xinyuan Wang, Xu Mou, and Yonghao Wang. 2023 · 2023
Cited alongside, same era.
# instag: Instruction tagging for analyzing supervised fine-tuning of large language models
Keming Lu, Hongyi Yuan, Zheng Yuan, Runji Lin, Junyang Lin, Chuanqi Tan, Chang Zhou, and Jingren Zhou. 2023 · 2023
Cited alongside, same era.
Automix: Automatically mixing language models
Aman Madaan, Pranjal Aggarwal, Ankit Anand, Srividya Pranavi Potharaju, Swaroop Mishra, Pei Zhou, Aditya Gupta, Dheeraj Rajagopal, Karthik Kappaganthu, Yiming Yang, et al. 2023 · 2023
Sade: A speaker-aware dual encoding model based on diagbert for medical triage and pre-diagnosis
Haozhou Li, Xinyuan Wang, Hongkai Du, Wentong Sun, and Qinke Peng. 2024 · 2024
Later among the works it cites.
MathBench: Evaluating the theory and application proficiency of LLMs with a hierarchical mathematics benchmark
Hongwei Liu, Zilong Zheng, Yuxuan Qiao, Haodong Duan, Zhiwei Fei, Fengzhe Zhou, Wenwei Zhang, Songyang Zhang, Dahua Lin, and Kai Chen. 2024b · 2024
Later among the works it cites.
Metallm: A high-performant and cost-efficient dynamic framework for wrapping llms
Quang H Nguyen, Duy C Hoang, Juliette Decugis, Saurav Manchanda, Nitesh V Chawla, and Khoa D Doan. 2024 · 2024
Later among the works it cites.
Routellm: Learning to route llms with preference data
Isaac Ong, Amjad Almahairi, Vincent Wu, Wei-Lin Chiang, Tianhao Wu, Joseph E Gonzalez, M Waleed Kadous, and Ion Stoica. 2024 · 2024
Later among the works it cites.
Observational scaling laws and the predictability of language model performance
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Large language model routing with benchmark datasets
Tal Shnitzer, Anthony Ou, Mírian Silva, Kate Soule, Yuekai Sun, Justin Solomon, Neil Thompson, and Mikhail Yurochkin. 2023 · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Cited alongside, same era.
Hybrid llm: Cost-efficient and quality-aware query routing
Dujian Ding, Ankur Mallick, Chi Wang, Robert Sim, Subhabrata Mukherjee, Victor Ruhle, Laks VS Lakshmanan, and Ahmed Hassan Awadallah. 2024 · 2024
Cited alongside, same era.
Evolutionary large language model for automated feature transformation
Nanxu Gong, Chandan K Reddy, Wangyang Ying, and Yanjie Fu. 2024 · 2024
Cited alongside, same era.
Empowering time series analysis with large language models: A survey
Yushan Jiang, Zijie Pan, Xikun Zhang, Sahil Garg, Anderson Schneider, Yuriy Nevmyvaka, and Dongjin Song. 2024 · 2024
Cited alongside, same era.
Routerbench: A benchmark for multi-llm routing system
Qitian Jason Hu, Jacob Bieker, Xiuyu Li, Nan Jiang, Benjamin Keigwin, Gaurav Ranganath, Kurt Keutzer, and Shriyash Kaustubh Upadhyay. 2024a
Cited in the paper.
Reinforcement feature transformation for polymer property performance prediction
Xuanming Hu, Dongjie Wang, Wangyang Ying, and Yanjie Fu. 2024b
Cited in the paper.
Yangjun Ruan, Chris J Maddison, and Tatsunori Hashimoto. 2024 · 2024
Later among the works it cites.
Fly-swat or cannon? cost-effective language model choice via meta-modeling
Marija Šakota, Maxime Peyrard, and Robert West. 2024 · 2024
Later among the works it cites.
Spatio-temporal transformer for temperature profiles prediction in large format additive manufacturing
Haoyang Xie, Dylan Hoskins, Kyle Rowe, and Feng Ju. 2024 · 2024
Later among the works it cites.
Revolutionizing biomarker discovery: Leveraging generative ai for bio-knowledge-embedded continuous space exploration
Wangyang Ying, Dongjie Wang, Xuanming Hu, Ji Qiu, Jin Park, and Yanjie Fu. 2024 · 2024
Later among the works it cites.
Calorie restriction induces mandible bone loss by regulating mitochondrial function
Linyi Liu, Phuong T Le, Victoria E DeMambro, Tiange Feng, Hanghang Liu, Wangyang Ying, Roland Baron, and Clifford J Rosen. 2025 · 2025
Closest in time.
Transformer-based offline printing strategy design for large format additive manufacturing
Haoyang Xie, Dylan Hoskins, Kyle Rowe, and Feng Ju. 2025 · 2025
Closest in time.