Fetching the paper…
Reading the bibliography…
Although the annotation paradigm based on Large Language Models (LLMs) has made significant breakthroughs in recent years, its actual deployment still has two core bottlenecks: first, the cost of calling commercial APIs in large-scale annotation is very expensive; second, in scenarios that require fine-grained semantic understanding, such as sentiment classification and toxicity classification, the annotation accuracy of LLMs is even lower than that of Small Language Models (SLMs) dedicated to this field.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Target-dependent twitter sentiment classification
Long Jiang, Mo Yu, Ming Zhou, Xiaohua Liu, and Tiejun Zhao. 2011 · 2011
Earlier work this paper cites.
Challenges for toxic comment classification: An in-depth error analysis
Betty Van Aken, Julian Risch, Ralf Krestel, and Alexander Löser. 2018 · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Remi Denton, Mark Díaz, Ian Kivlichan, Vinodkumar Prabhakaran, and Rachel Rosen. 2021 · 2021
Earlier work this paper cites.
Zeyad Emam, Andrew Kondrich, Sasha Harrison, Felix Lau, Yushi Wang, Aerin Kim, and Elliot Branson. 2021 · 2021
Earlier work this paper cites.
A survey on aspect-based sentiment classification
Gianni Brauwers and Flavius Frasincar. 2022 · 2022
Earlier work this paper cites.
The challenge of data annotation in deep learning—a case study on whole plant corn silage
Christoffer Bøgelund Rasmussen, Kristian Kirk, and Thomas B Moeslund. 2022 · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022 · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Earlier work this paper cites.
Can large language models fix data annotation errors? an empirical study using debatepedia for query-focused text summarization
Md Tahmid Rahman Laskar, Mizanur Rahman, Israt Jahan, Enamul Hoque, and Jimmy Huang. 2023 · 2023
Earlier work this paper cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Romain Sauvestre, Tal Remez, et al. 2023 · 2023
Earlier work this paper cites.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto. 2023 · 2023
Earlier work this paper cites.
Leveraging large language models and weak supervision for social media data annotation: an evaluation using covid-19 self-reported vaccination tweets
Ramya Tekumalla and Juan M Banda. 2023 · 2023
Earlier work this paper cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Earlier work this paper cites.
A review on code generation with llms: Application and evaluation
Jianxun Wang and Yixiang Chen. 2023 · 2023
Cited alongside, same era.
Next-gpt: Any-to-any multimodal llm
Shengqiong Wu, Hao Fei, Leigang Qu, Wei Ji, and Tat-Seng Chua. 2023 · 2023
Cited alongside, same era.
Small models are valuable plug-ins for large language models
Canwen Xu, Yichong Xu, Shuohang Wang, Yang Liu, Chenguang Zhu, and Julian McAuley. 2023 · 2023
Cited alongside, same era.
Baichuan 2: Open large-scale language models
Aiyuan Yang, Bin Xiao, Bingning Wang, Borong Zhang, Ce Bian, Chao Yin, Chenxu Lv, Da Pan, Dian Wang, Dong Yan, et al. 2023 · 2023
Cited alongside, same era.
Fill in the gaps: Model calibration and generalization with synthetic data
Determlr: Augmenting llm-based logical reasoning from indeterminacy to determinacy
Hongda Sun, Weikai Xu, Wei Liu, Jian Luan, Bin Wang, Shuo Shang, Ji-Rong Wen, and Rui Yan. 2024 · 2024
Later among the works it cites.
Enhancing text annotation through rationale-driven collaborative few-shot prompting
Jianfei Wu, Xubin Wang, and Weijia Jia. 2024 · 2024
Later among the works it cites.
Towards automating text annotation: A case study on semantic proximity annotation using gpt-4
Sachin Yadav, Tejaswi Choppa, and Dominik Schlechtweg. 2024 · 2024
Later among the works it cites.
An Yang, Baosong Yang, Beichen Zhang, Binyuan Hui, Bo Zheng, Bowen Yu, Chengyuan Li, Dayiheng Liu, Fei Huang, Haoran Wei, et al. 2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yang Ba, Michelle V Mancenido, and Rong Pan. 2024 · 2024
Cited alongside, same era.
Is a large language model a good annotator for event extraction?
Ruirui Chen, Chengwei Qin, Weifeng Jiang, and Dongkyu Choi. 2024 · 2024
Cited alongside, same era.
Multi-news+: Cost-efficient dataset cleansing via llm-based data annotation
Juhwan Choi, Jungmin Yun, Kyohoon Jin, and YoungBin Kim. 2024 · 2024
Cited alongside, same era.
Large language models improve annotation of prokaryotic viral proteins
Zachary N Flamholz, Steven J Biller, and Libusha Kelly. 2024 · 2024
Cited alongside, same era.
Aaron Grattafiori, Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, Abhishek Kadian, Ahmad Al-Dahle, Aiesha Letman, Akhil Mathur, Alan Schelten, Alex Vaughan, et al. 2024 · 2024
Cited alongside, same era.
You only prompt once: On the capabilities of prompt learning on large language models to tackle toxic content
Xinlei He, Savvas Zannettou, Yun Shen, and Yang Zhang. 2024 · 2024
Cited alongside, same era.
On limitations of llm as annotator for low resource languages
Suramya Jadhav, Abhay Shanbhag, Amogh Thakurdesai, Ridhima Sinare, and Raviraj Joshi. 2024 · 2024
Cited alongside, same era.
A survey on large language models for code generation
Juyong Jiang, Fan Wang, Jiasi Shen, Sungju Kim, and Sunghun Kim. 2024 · 2024
Cited alongside, same era.
Kaiyan Zhang, Jianyu Wang, Ermo Hua, Biqing Qi, Ning Ding, and Bowen Zhou. 2024 · 2024
Later among the works it cites.
Llm-empowered embodied agent for memory-augmented task planning in household robotics
Marc Glocker, Peter Hönig, Matthias Hirschmanner, and Markus Vincze. 2025 · 2025
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al. 2025 · 2025
Closest in time.
A zero-shot generalization framework for llm-driven cross-domain sequential recommendation
Yunzhe Li, Junting Wang, Hari Sundaram, and Zhining Liu. 2025 · 2025
Closest in time.
Xiao Liu, Zirui Wu, Jiayi Li, Zhicheng Shao, Xun Pang, and Yansong Feng. 2025 · 2025
Closest in time.
Kimi k1. 5: Scaling reinforcement learning with llms
Kimi Team, Angang Du, Bofei Gao, Bowei Xing, Changjiu Jiang, Cheng Chen, Cheng Li, Chenjun Xiao, Chenzhuang Du, Chonghua Liao, et al. 2025 · 2025
Closest in time.
Ran Xu, Wenqi Shi, Yuchen Zhuang, Yue Yu, Joyce C Ho, Haoyu Wang, and Carl Yang. 2025 · 2025
Closest in time.
Limo: Less is more for reasoning
Yixin Ye, Zhen Huang, Yang Xiao, Ethan Chern, Shijie Xia, and Pengfei Liu. 2025 · 2025
Closest in time.
Citer: Collaborative inference for efficient large language model decoding with token-level routing
Wenhao Zheng, Yixiao Chen, Weitong Zhang, Souvik Kundu, Yun Li, Zhengzhong Liu, Eric P Xing, Hongyi Wang, and Huaxiu Yao. 2025 · 2025
Closest in time.
Optimizing code retrieval: High-quality and scalable dataset annotation through large language models
Rui Li, Qi Liu, Liyang He, Zheng Zhang, Hao Zhang, Shengyu Ye, Junyu Lu, and Zhenya Huang. 2024b · 2065
Closest in time.