Fetching the paper…
Reading the bibliography…
Large-scale deployment of generative AI tools often depends on costly API calls to a Large Language Model (LLM) to fulfil user queries.
Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning
Haokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin A Raffel. 2022 · 1965
Earlier work this paper cites.
Query by committee
H. S. Seung, M. Opper, and H. Sompolinsky. 1992 · 1992
Earlier work this paper cites.
Less is more: Active learning with support vector machines
Greg Schohn and David Cohn. 2000 · 2000
Earlier work this paper cites.
Toward optimal active learning through sampling estimation of error reduction
Nicholas Roy and Andrew McCallum. 2001 · 2001
Earlier work this paper cites.
Active Hidden Markov Models for Information Extraction
Tobias Scheffer, Christian Decomain, and Stefan Wrobel. 2001 · 2001
Earlier work this paper cites.
Active learning to recognize multiple types of plankton
Tong Luo, K. Kramer, S. Samson, A. Remsen, D.B. Goldgof, L.O. Hall, and T. Hopkins. 2004 · 2004
Earlier work this paper cites.
Seeing Stars: Exploiting Class Relationships for Sentiment Categorization with Respect to Rating Scales
Bo Pang and Lillian Lee. 2005 · 2005
Earlier work this paper cites.
Model compression
Cristian Bucila, Rich Caruana, and Alexandru Niculescu-Mizil. 2006 · 2006
Earlier work this paper cites.
Evaluation of text generation: A survey
Asli Celikyilmaz, Elizabeth Clark, and Jianfeng Gao. 2020 · 2006
Earlier work this paper cites.
MMR-based Active Machine Learning for Bio Named Entity Recognition
Seokhwan Kim, Yu Song, Kyungduk Kim, Jeong-Won Cha, and Gary Geunbae Lee. 2006 · 2006
Earlier work this paper cites.
Margin-based active learning for structured output spaces
Dan Roth and Kevin Small. 2006 · 2006
Earlier work this paper cites.
Margin Based Active Learning
Maria-Florina Balcan, Andrei Broder, and Tong Zhang. 2007 · 2007
Earlier work this paper cites.
Learning from time-changing data with adaptive windowing
Albert Bifet and Ricard Gavaldà. 2007 · 2007
Earlier work this paper cites.
Active Learning for Regression Based on Query by Committee
Robert Burbidge, Jem J. Rowland, and Ross D. King. 2007 · 2007
Earlier work this paper cites.
Active Learning Literature Survey
Burr Settles. 2009 · 2012
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeffrey Dean. 2015 · 2015
Cited alongside, same era.
Universality Versus Cultural Specificity of Three Emotion Domains: Some Evidence Based on the Cascading Model of Emotional Intelligence
Bo Shao, Lorna Doucet, and David R. Caruso. 2015 · 2015
Cited alongside, same era.
Online Decision Making for Statistical Model Fitting
Carlos Riquelme. 2017 · 2017
Cited alongside, same era.
Can a suit of armor conduct electricity? A new dataset for open book question answering
Todor Mihaylov, Peter Clark, Tushar Khot, and Ashish Sabharwal. 2018 · 2018
Cited alongside, same era.
Active learning for convolutional neural networks: A core-set approach
Ozan Sener and Silvio Savarese. 2018 · 2018
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
Sentence-t5: Scalable sentence encoders from pre-trained text-to-text models
Jianmo Ni, Gustavo Hernández Ábrego, Noah Constant, Ji Ma, Keith B. Hall, Daniel Cer, and Yinfei Yang. 2022 · 2022
Later among the works it cites.
A survey of deep active learning
Pengzhen Ren, Yun Xiao, Xiaojun Chang, Po-Yao Huang, Zhihui Li, Brij B. Gupta, Xiaojiang Chen, and Xin Wang. 2022 · 2022
Later among the works it cites.
Revisiting Uncertainty-based Query Strategies for Active Learning with Transformers
Christopher Schröder, Andreas Niekler, and Martin Potthast. 2022 · 2022
Later among the works it cites.
A comparative survey of deep active learning
Xueying Zhan, Qingzhong Wang, Kuan-Hao Huang, Haoyi Xiong, Dejing Dou, and Antoni B. Chan. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FEVER: a large-scale dataset for fact extraction and verification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Empirical Evaluation of Active Learning Techniques for Neural MT
Xiangkai Zeng, Sarthak Garg, Rajen Chatterjee, Udhyakumar Nallasamy, and Matthias Paulik. 2019 · 2019
Cited alongside, same era.
Pseudo-labeling and confirmation bias in deep semi-supervised learning
Eric Arazo, Diego Ortego, Paul Albert, Noel E O’Connor, and Kevin McGuinness. 2020 · 2020
Cited alongside, same era.
Continual lifelong learning in natural language processing: A survey
Magdalena Biesialska, Katarzyna Biesialska, and Marta R. Costa-jussà. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
Green AI
Roy Schwartz, Jesse Dodge, Noah A. Smith, and Oren Etzioni. 2020 · 2020
Cited alongside, same era.
A survey of active learning for natural language processing
Zhisong Zhang, Emma Strubell, and Eduard H. Hovy. 2022 · 2022
Later among the works it cites.
Gptcache
Zilliz. 2023 · 2022
Later among the works it cites.
Robust active distillation
Cenk Baykal, Khoa Trinh, Fotis Iliopoulos, Gaurav Menghani, and Erik Vee. 2023 · 2023
Closest in time.
A survey on online active learning
Davide Cacciarelli and Murat Kulahci. 2023 · 2023
Closest in time.
Frugalgpt: How to use large language models while reducing cost and improving performance
Lingjiao Chen, Matei Zaharia, and James Zou. 2023 · 2023
Closest in time.
Combining parameter-efficient modules for task-level generalisation
Edoardo Maria Ponti, Alessandro Sordoni, Yoshua Bengio, and Siva Reddy. 2023 · 2023
Closest in time.
Large language model routing with benchmark datasets
Tal Shnitzer, Anthony Ou, Mírian Silva, Kate Soule, Yuekai Sun, Justin Solomon, Neil Thompson, and Mikhail Yurochkin. 2023 · 2023
Closest in time.
Fly-swat or cannon? cost-effective language model choice via meta-modeling
Marija Šakota, Maxime Peyrard, and Robert West. 2023 · 2023
Closest in time.
Computation-efficient knowledge distillation via uncertainty-aware mixup
Guodong Xu, Ziwei Liu, and Chen Change Loy. 2023 · 2023
Closest in time.
On optimal caching and model multiplexing for large model inference
Banghua Zhu, Ying Sheng, Lianmin Zheng, Clark W. Barrett, Michael I. Jordan, and Jiantao Jiao. 2023 · 2023
Closest in time.