Fetching the paper…
Reading the bibliography…
While large language models (LLMs) like ChatGPT have shown impressive capabilities in Natural Language Processing (NLP) tasks, a systematic investigation of their potential in this field remains largely unexplored.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Sentiment analysis: mining opinions, sentiments, and emotions, 2016
Jun Zhao, Kang Liu, and Liheng Xu · 2016
Earlier work this paper cites.
An overview of end-to-end language understanding and dialog management for personal digital assistants
Ruhi Sarikaya, Paul A Crook, Alex Marin, Minwoo Jeong, Jean-Philippe Robichaud, Asli Celikyilmaz, Young-Bum Kim, Alexandre Rochette, Omar Zia Khan, Xiaohu Liu, et al · 2016
Earlier work this paper cites.
Neural abstractive text summarization with sequence-to-sequence models
Tian Shi, Yaser Keneshloo, Naren Ramakrishnan, and Chandan K. Reddy · 2018
Earlier work this paper cites.
Parameter-efficient transfer learning for nlp
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly · 2019
Earlier work this paper cites.
Multimodal representation learning for recommendation in internet of things
Zhenhua Huang, Xin Xu, Juan Ni, Honghao Zhu, and Cheng Wang · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, Weizhu Chen, et al · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al · 2021
Earlier work this paper cites.
Yue Wang, Weishi Wang, Shafiq Joty, and Steven CH Hoi · 2021
Earlier work this paper cites.
A survey on spoken language understanding: Recent advances and new frontiers
Libo Qin, Tianbao Xie, Wanxiang Che, and Ting Liu · 2021
Earlier work this paper cites.
Codexglue: A machine learning benchmark dataset for code understanding and generation
Shuai Lu, Daya Guo, Shuo Ren, Junjie Huang, Alexey Svyatkovskiy, Ambrosio Blanco, Colin Clement, Dawn Drain, Daxin Jiang, Duyu Tang, et al · 2021
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel · 2021
Earlier work this paper cites.
Language models are few-shot multilingual learners
Genta Indra Winata, Andrea Madotto, Zhaojiang Lin, Rosanne Liu, Jason Yosinski, and Pascale Fung · 2021
Earlier work this paper cites.
Bold: Dataset and metrics for measuring biases in open-ended language generation
J. Dhamala, Tony Sun, Varun Kumar, Satyapriya Krishna, Yada Pruksachatkun, Kai-Wei Chang, and Rahul Gupta · 2021
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Earlier work this paper cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Earlier work this paper cites.
Rethinking the role of demonstrations: What makes in-context learning work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Earlier work this paper cites.
A survey on sentiment analysis methods, applications, and challenges
Mayur Wankhade, Annavarapu Chandra Sekhara Rao, and Chaitanya Kulkarni · 2022
Earlier work this paper cites.
Controllable dialogue simulation with in-context learning
Zekun Li, Wenhu Chen, Shiyang Li, Hong Wang, Jing Qian, and Xifeng Yan · 2022
Earlier work this paper cites.
In-context learning for few-shot dialogue state tracking
Yushi Hu, Chia-Hsuan Lee, Tianbao Xie, Tao Yu, Noah A Smith, and Mari Ostendorf · 2022
Earlier work this paper cites.
Binding language models in symbolic languages
Zhoujun Cheng, Tianbao Xie, Peng Shi, Chengzu Li, Rahul Nadkarni, Yushi Hu, Caiming Xiong, Dragomir Radev, Mari Ostendorf, Luke Zettlemoyer, et al · 2022
Earlier work this paper cites.
News summarization and evaluation in the era of gpt-3
Tanya Goyal, Junyi Jessy Li, and Greg Durrett · 2022
Earlier work this paper cites.
Prompted opinion summarization with gpt-3.5
Adithya Bhaskar, Alexander R. Fabbri, and Greg Durrett · 2022
Earlier work this paper cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong · 2022
Earlier work this paper cites.
Pangu-coder: Program synthesis with function-level language modeling
Fenia Christopoulou, Gerasimos Lampouras, Milan Gritta, Guchun Zhang, Yinpeng Guo, Zhongqi Li, Qi Zhang, Meng Xiao, Bo Shen, Lin Li, et al · 2022
Earlier work this paper cites.
Instruction tuning for few-shot aspect-based sentiment analysis
Siddharth Varia, Shuai Wang, Kishaloy Halder, Robert Vacareanu, Miguel Ballesteros, Yassine Benajiba, Neha Anna John, Rishita Anubhai, Smaranda Muresan, and Dan Roth · 2022
Earlier work this paper cites.
Unifiedskg: Unifying and multi-tasking structured knowledge grounding with text-to-text language models
Tianbao Xie, Chen Henry Wu, Peng Shi, Ruiqi Zhong, Torsten Scholak, Michihiro Yasunaga, Chien-Sheng Wu, Ming Zhong, Pengcheng Yin, Sida I Wang, et al · 2022
Earlier work this paper cites.
Show, don’t tell: Demonstrations outperform descriptions for schema-guided task-oriented dialogue
Raghav Gupta, Harrison Lee, Jeffrey Zhao, Yuan Cao, Abhinav Rastogi, and Yonghui Wu · 2022
Earlier work this paper cites.
Knowledge-grounded dialog state tracking
Dian Yu, Mingqiu Wang, Yuan Cao, Laurent El Shafey, Izhak Shafran, and Hagen Soltau · 2022
Earlier work this paper cites.
Socratic pretraining: Question-driven pretraining for controllable summarization
Artidoro Pagnoni, Alexander R. Fabbri, Wojciech Kryscinski, and Chien-Sheng Wu · 2022
Earlier work this paper cites.
Domain-oriented prefix-tuning: Towards efficient and generalizable fine-tuning for zero-shot dialogue summarization
Lulu Zhao, Fujia Zheng, Weihao Zeng, Keqing He, Weiran Xu, Huixing Jiang, Wei Wu, and Yanan Wu · 2022
Earlier work this paper cites.
Few-shot query-focused summarization with prefix-merging
Ruifeng Yuan, Zili Wang, Ziqiang Cao, and Wenjie Li · 2022
Earlier work this paper cites.
Coderl: Mastering code generation through pretrained models and deep reinforcement learning
Hung Le, Yue Wang, Akhilesh Deepak Gotmare, Silvio Savarese, and Steven Chu Hong Hoi · 2022
Earlier work this paper cites.
Parameter-efficient finetuning of transformers for source code
Shamil Ayupov and Nadezhda Chirkova · 2022
Earlier work this paper cites.
When does parameter-efficient transfer learning work for machine translation?
A. Ustun and Asa Cooper Stickland · 2022
Earlier work this paper cites.
A survey on table question answering: recent advances
Nengzheng Jin, Joanna Siebert, Dongfang Li, and Qingcai Chen · 2022
Earlier work this paper cites.
Bloom: A 176b-parameter open-access multilingual language model
BigScience Workshop, Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, et al · 2022
Earlier work this paper cites.
Language models are multilingual chain-of-thought reasoners
Freda Shi, Mirac Suzgun, Markus Freitag, Xuezhi Wang, Suraj Srivats, Soroush Vosoughi, Hyung Won Chung, Yi Tay, Sebastian Ruder, Denny Zhou, et al · 2022
Earlier work this paper cites.
Few-shot learning with multilingual generative language models
Xi Victoria Lin, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, Ramakanth Pasunuru, Sam Shleifer, Punit Singh Koura, Vishrav Chaudhary, Brian O’Horo, Jeff Wang, Luke Zettlemoyer, Zornitsa Kozareva, Mona Diab, Veselin Stoyanov, and Xian Li · 2022
Earlier work this paper cites.
Learn to explain: Multimodal reasoning via thought chains for science question answering
Pan Lu, Swaroop Mishra, Tanglin Xia, Liang Qiu, Kai-Wei Chang, Song-Chun Zhu, Oyvind Tafjord, Peter Clark, and Ashwin Kalyan · 2022
Earlier work this paper cites.
Wenhu Chen, Xueguang Ma, Xinyi Wang, and William W Cohen · 2022
Earlier work this paper cites.
A token-level reference-free hallucination detection benchmark for free-form text generation
Tianyu Liu, Yizhe Zhang, Chris Brockett, Yi Mao, Zhifang Sui, Weizhu Chen, and William B Dolan · 2022
Earlier work this paper cites.
ToxiGen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection
Thomas Hartvigsen, Saadia Gabriel, Hamid Palangi, Maarten Sap, Dipankar Ray, and Ece Kamar · 2022
Earlier work this paper cites.
Red teaming language models to reduce harms: Methods, scaling behaviors, and lessons learned
Deep Ganguli, Liane Lovitt, Jackson Kernion, Amanda Askell, Yuntao Bai, Saurav Kadavath, Ben Mann, Ethan Perez, Nicholas Schiefer, Kamal Ndousse, et al · 2022
Earlier work this paper cites.
Challenges and applications of large language models
Jean Kaddour, Joshua Harris, Maximilian Mozes, Herbie Bradley, Roberta Raileanu, and Robert McHardy · 2023
Earlier work this paper cites.
Large language models: a comprehensive survey of its applications, challenges, limitations, and future prospects
Muhammad Usman Hadi, Rizwan Qureshi, Abbas Shah, Muhammad Irfan, Anas Zafar, Muhammad Bilal Shaikh, Naveed Akhtar, Jia Wu, Seyedali Mirjalili, et al · 2023
Earlier work this paper cites.
Through the lens of core competency: Survey on evaluation of large language models
Ziyu Zhuang, Qiguang Chen, Longxuan Ma, Mingda Li, Yi Han, Yushan Qian, Haopeng Bai, Zixian Feng, Weinan Zhang, and Ting Liu · 2023
Earlier work this paper cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Earlier work this paper cites.
Towards making the most of chatgpt for machine translation
Keqin Peng, Liang Ding, Qihuang Zhong, Li Shen, Xuebo Liu, Min Zhang, Yuanxin Ouyang, and Dacheng Tao · 2023
Earlier work this paper cites.
Qlora: Efficient finetuning of quantized llms
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer · 2023
Earlier work this paper cites.
Beyond information: Is chatgpt empathetic enough?
Ahmed Belkhir and Fatiha Sadat · 2023
Earlier work this paper cites.
Empirical study of zero-shot NER with ChatGPT
Tingyu Xie, Qi Li, Jian Zhang, Yan Zhang, Zuozhu Liu, and Hongwei Wang · 2023
Earlier work this paper cites.
How far is language model from 100% few-shot named entity recognition in medical domain
Mingchen Li and Rui Zhang · 2023
Earlier work this paper cites.
Codekgc: Code language model for generative knowledge graph construction
Zhen Bi, Jing Chen, Yinuo Jiang, Feiyu Xiong, Wei Guo, Huajun Chen, and Ningyu Zhang · 2023
Earlier work this paper cites.
A preliminary evaluation of chatgpt for zero-shot dialogue understanding
Wenbo Pan, Qiguang Chen, Xiao Xu, Wanxiang Che, and Libo Qin · 2023
Earlier work this paper cites.
Can chatgpt detect intent? evaluating large language models for spoken language understanding
Mutian He and Philip N Garner · 2023
Earlier work this paper cites.
Are llms all you need for task-oriented dialogue?
Vojtěch Hudeček and Ondřej Dušek · 2023
Earlier work this paper cites.
Chatgpt for zero-shot dialogue state tracking: A solution or an opportunity?
Michael Heck, Nurul Lubis, Benjamin Ruppik, Renato Vukovic, Shutong Feng, Christian Geishauser, Hsien-Chin Lin, Carel van Niekerk, and Milica Gašić · 2023
Earlier work this paper cites.
Dialogue distillery: Crafting interpolable, interpretable, and introspectable dialogue from llms
Ryan A Chi, Jeremy Kim, Scott Hickmann, Siyan Li, Gordon Chi, Thanawan Atchariyachanvanit, Katherine Yu, Nathan A Chi, Gary Dai, Shashank Rammoorthy, et al · 2023
Earlier work this paper cites.
Diverse retrieval-augmented in-context learning for dialogue state tracking
Brendan King and Jeffrey Flanigan · 2023
Earlier work this paper cites.
Multi-party goal tracking with llms: Comparing pre-training, fine-tuning, and prompt engineering
Angus Addlesee, Weronika Sieińska, Nancie Gunson, Daniel Hernández Garcia, Christian Dondrup, and Oliver Lemon · 2023
Earlier work this paper cites.
Instructtods: Large language models for end-to-end task-oriented dialogue systems
Willy Chung, Samuel Cahyawijaya, Bryan Wilie, Holy Lovenia, and Pascale Fung · 2023
Earlier work this paper cites.
Orchestrallm: Efficient orchestration of language models for dialogue state tracking
Chia-Hsuan Lee, Hao Cheng, and Mari Ostendorf · 2023
Cited alongside, same era.
Toward a better understanding of the emotional dynamics of negotiation with large language models
Eleanor Lin, James Hale, and Jonathan Gratch · 2023
Cited alongside, same era.
Diaggpt: An llm-based chatbot with automatic topic management for task-oriented dialogue
Lang Cao · 2023
Cited alongside, same era.
Tabular representation, noisy operators, and impacts on table structure understanding tasks in llms
Ananya Singha, José Cambronero, Sumit Gulwani, Vu Le, and Chris Parnin · 2023
Cited alongside, same era.
Large language models are versatile decomposers: Decomposing evidence and questions for table-based reasoning
Yunhu Ye, Binyuan Hui, Min Yang, Binhua Li, Fei Huang, and Yongbin Li · 2023
Structured information extraction from scientific text with large language models
John Dagdelen, Alexander Dunn, Sanghoon Lee, Nicholas Walker, Andrew S Rosen, Gerbrand Ceder, Kristin A Persson, and Anubhav Jain · 2024
Closest in time.
Autore: Document-level relation extraction with large language models
Lilong Xue, Dan Zhang, Yuxiao Dong, and Jie Tang · 2024
Closest in time.
Interleaved multi-modal document representations for large-scale information retrieval using large language models
Dominic Rixewa, Katherine Anderson, Leonard Dubois, and Mateo Harrington · 2024
Closest in time.
Table-gpt: Table fine-tuned gpt for diverse table tasks
Peng Li, Yeye He, Dror Yashar, Weiwei Cui, Song Ge, Haidong Zhang, Danielle Rifinski Fainman, Dongmei Zhang, and Surajit Chaudhuri · 2024
Closest in time.
Astraios: Parameter-efficient instruction tuning code large language models
Terry Yue Zhuo, Armel Zebaze, Nitchakarn Suppattarachai, Leandro von Werra, Harm de Vries, Qian Liu, and Niklas Muennighoff · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
CRT-QA: A dataset of complex reasoning question answering over tabular data
Zhehao Zhang, Xitao Li, Yan Gao, and Jian-Guang Lou · 2023
Cited alongside, same era.
Large language models are few (1)-shot table reasoners
Wenhu Chen · 2023
Cited alongside, same era.
StructGPT: A general framework for large language model to reason over structured data
Jinhao Jiang, Kun Zhou, Zican Dong, Keming Ye, Xin Zhao, and Ji-Rong Wen · 2023
Cited alongside, same era.
Zero-shot cross-lingual summarization via large language models
Jiaan Wang, Yunlong Liang, Fandong Meng, Beiqi Zou, Zhixu Li, Jianfeng Qu, and Jie Zhou · 2023
Cited alongside, same era.
From sparse to dense: Gpt-4 summarization with chain of density prompting
Griffin Adams, Alexander R. Fabbri, Faisal Ladhak, Eric Lehman, and Noémie Elhadad · 2023
Cited alongside, same era.
In-context learning of large language models for controlled dialogue summarization: A holistic benchmark and empirical analysis
Yuting Tang, Ratish Puduppully, Zhengyuan Liu, and Nancy Chen · 2023
Cited alongside, same era.
Santacoder: don’t reach for the stars!
Loubna Ben Allal, Raymond Li, Denis Kocetkov, Chenghao Mou, Christopher Akiki, Carlos Munoz Ferrandis, Niklas Muennighoff, Mayank Mishra, Alex Gu, Manan Dey, et al · 2023
Cited alongside, same era.
Closest in time.
Haoran Xu, Amr Sharaf, Yunmo Chen, Weiting Tan, Lingfeng Shen, Benjamin Van Durme, Kenton Murray, and Young Jin Kim · 2024
Closest in time.
Math-llava: Bootstrapping mathematical reasoning for multimodal large language models
Wenhao Shi, Zhiqiang Hu, Yi Bin, Junhua Liu, Yang Yang, See-Kiong Ng, Lidong Bing, and Roy Ka-Wei Lee · 2024
Closest in time.
Deepseekmath: Pushing the limits of mathematical reasoning in open language models
Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, Haowei Zhang, Mingchuan Zhang, YK Li, Y Wu, et al · 2024
Closest in time.
Improve mathematical reasoning in language models by automated process supervision
Liangchen Luo, Yinxiao Liu, Rosanne Liu, Samrat Phatale, Meiqi Guo, Harsh Lara, Yunxuan Li, Lei Shu, Yun Zhu, Lei Meng, et al · 2024
Closest in time.
Reasonagain: Using extractable symbolic programs to evaluate mathematical reasoning
Xiaodong Yu, Ben Zhou, Hao Cheng, and Dan Roth · 2024
Closest in time.
System-2 mathematical reasoning via enriched instruction tuning
Huanqia Cai, Yijun Yang, and Zhifeng Li · 2024
Closest in time.
Blendx: Complex multi-intent detection with blended patterns
Yejin Yoon, Jungyeon Lee, Kangsan Kim, Chanhee Park, and Taeuk Kim · 2024
Closest in time.
Qwen2.5-coder technical report
Binyuan Hui, Jian Yang, Zeyu Cui, Jiaxi Yang, Dayiheng Liu, Lei Zhang, Tianyu Liu, Jiajun Zhang, Bowen Yu, Keming Lu, et al · 2024
Closest in time.
André Storhaug and Jingyue Li · 2024
Closest in time.
Aligning translation-specific understanding to general understanding in large language models
Yi-Chong Huang, Xiaocheng Feng, Baohang Li, Chengpeng Fu, Wenshuai Huo, Ting Liu, and Bing Qin · 2024
Closest in time.
Towards robust in-context learning for machine translation with large language models
Shaolin Zhu, Menglong Cui, and Deyi Xiong · 2024
Closest in time.
Ladder: A model-agnostic framework boosting llm-based machine translation to the next level
Zhaopeng Feng, Ruizhe Chen, Yan Zhang, Zijie Meng, and Zuozhu Liu · 2024
Closest in time.
Parameter-efficient fine-tuning of large language models using semantic knowledge tuning
Nusrat Jahan Prottasha, Asif Mahmud, Md Shohanur Islam Sobuj, Prakash Bhat, Md Kowsher, Niloofar Yousefi, and Ozlem Ozmen Garibay · 2024
Closest in time.
Video-of-thought: step-by-step video reasoning from perception to cognition
Hao Fei, Shengqiong Wu, Wei Ji, Hanwang Zhang, Meishan Zhang, Mong Li Lee, and Wynne Hsu · 2024
Closest in time.
Mengkang Hu, Tianxing Chen, Qiguang Chen, Yao Mu, Wenqi Shao, and Ping Luo · 2024
Closest in time.
Wrong-of-thought: An integrated reasoning framework with multi-perspective verification and wrong information
Yongheng Zhang, Qiguang Chen, Jingxuan Zhou, Peng Wang, Jiasheng Si, Jin Wang, Wenpeng Lu, and Libo Qin · 2024
Closest in time.
Aaron Jaech, Adam Kalai, Adam Lerer, Adam Richardson, Ahmed El-Kishky, Aiden Low, Alec Helyar, Aleksander Madry, Alex Beutel, Alex Carney, et al · 2024
Closest in time.
Self-reflection in llm agents: Effects on problem-solving performance
Matthew Renze and Erhan Guven · 2024
Closest in time.
Refiner: Reasoning feedback on intermediate representations
Debjit Paul, Mete Ismayilzada, Maxime Peyrard, Beatriz Borges, Antoine Bosselut, Robert West, and Boi Faltings · 2024
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al · 2025
Closest in time.
Dylas: A dynamic label alignment strategy for large-scale multi-label text classification
Lin Ren, Yongbin Liu, Chunping Ouyang, Ying Yu, Shuda Zhou, Yidong He, and Yaping Wan · 2025
Closest in time.
Tablelora: Low-rank adaptation on table structure understanding for large language models
Xinyi He, Yihao Liu, Mengyu Zhou, Yeye He, Haoyu Dong, Shi Han, Zejian Yuan, and Dongmei Zhang · 2025
Closest in time.
Improving chain-of-thought reasoning via quasi-symbolic abstractions
Leonardo Ranaldi, Marco Valentino, Alexander Polonsky, and Andrè Freitas · 2025
Closest in time.
Debate, train, evolve: Self evolution of language model reasoning
Gaurav Srivastava, Zhenyu Bi, Meng Lu, and Xuan Wang · 2025
Closest in time.
An automated information extraction model for unstructured discharge letters using large language models and gpt-4
Robert M Siepmann, Giulia Baldini, Cynthia S Schmidt, Daniel Truhn, Gustav Anton Müller-Franzes, Amin Dada, Jens Kleesiek, Felix Nensa, and René Hosch · 2025
Closest in time.
Scalable information extraction from free text electronic health records using large language models
Bowen Gu, Vivian Shao, Ziqian Liao, Valentina Carducci, Santiago Romero Brufau, Jie Yang, and Rishi J Desai · 2025
Closest in time.
Croprompt: Cross-task interactive prompting for zero-shot spoken language understanding
Libo Qin, Fuxuan Wei, Qiguang Chen, Jingxuan Zhou, Shijue Huang, Jiasheng Si, Wenpeng Lu, and Wanxiang Che · 2025
Closest in time.
ProTOD: Proactive task-oriented dialogue system based on large language model
Wenjie Dong, Sirong Chen, and Yan Yang · 2025
Closest in time.
Emre Can Acikgoz, Jeremiah Greer, Akul Datta, Ze Yang, William Zeng, Oussama Elachqar, Emmanouil Koukoumidis, Dilek Hakkani-Tür, and Gokhan Tur · 2025
Closest in time.
MIDLM: Multi-intent detection with bidirectional large language models
Shangjian Yin, Peijie Huang, and Yuhong Xu · 2025
Closest in time.
Leveraging long-context large language models for multi-document understanding and summarization in enterprise applications
Aditi Godbole, Jabin Geevarghese George, and Smita Shandilya · 2025
Closest in time.
Generalization bias in large language model summarization of scientific research
Uwe Peters and Benjamin Chin-Yee · 2025
Closest in time.
Summpilot: Bridging efficiency and customization for interactive summarization system
JungMin Yun, Juhwan Choi, Kyohoon Jin, Soojin Jang, Jinhee Jang, and YoungBin Kim · 2025
Closest in time.
Just what you desire: Constrained timeline summarization with self-reflection for enhanced relevance
Muhammad Reza Qorib, Qisheng Hu, and Hwee Tou Ng · 2025
Closest in time.
Dong-Hai Zhu, Yu-Jie Xiong, Jia-Chen Zhang, Xi-Jiong Xie, and Chun-Ming Xia · 2025
Closest in time.
Language models of code are few-shot planners and reasoners for multi-document summarization with attribution
Abhilash Nandy and Sambaran Bandyopadhyay · 2025
Closest in time.
Mutual reinforcement of llm dialogue synthesis and summarization capabilities for few-shot dialogue summarization
Yen-Ju Lu, Ting-Yao Hu, Hema Swetha Koppula, Hadi Pouransari, Jen-Hao Rick Chang, Yin Xia, Xiang Kong, Qi Zhu, Xiaoming Simon Wang, Oncel Tuzel, et al · 2025
Closest in time.
A dataset and benchmark for hospital course summarization with adapted large language models
Asad Aali, Dave Van Veen, Yamin Ishraq Arefeen, Jason Hom, Christian Bluethgen, Eduardo Pontes Reis, Sergios Gatidis, Namuun Clifford, Joseph Daws, Arash S Tehrani, et al · 2025
Closest in time.
Rlpf: Reinforcement learning from prediction feedback for user summarization with llms
Jiaxing Wu, Lin Ning, Luyang Liu, Harrison Lee, Neo Wu, Chao Wang, Sushant Prakash, Shawn O’Banion, Bradley Green, and Jun Xie · 2025
Closest in time.
Seed-coder: Let the code model curate data for itself
ByteDance Seed · 2025
Closest in time.
Ling Team, Wenting Cai, Yuchen Cao, Chaoyu Chen, Chen Chen, Siba Chen, Qing Cui, Peng Di, Junpeng Fang, Zi Gong, et al · 2025
Closest in time.
Process-supervised reinforcement learning for code generation
Yufan Ye, Ting Zhang, Wenbin Jiang, and Hua Huang · 2025
Closest in time.
Swe-rl: Advancing llm reasoning via reinforcement learning on open software evolution
Yuxiang Wei, Olivier Duchenne, Jade Copet, Quentin Carbonneaux, Lingming Zhang, Daniel Fried, Gabriel Synnaeve, Rishabh Singh, and Sida I Wang · 2025
Closest in time.
Yibo Yan, Shen Wang, Jiahao Huo, Philip S Yu, Xuming Hu, and Qingsong Wen · 2025
Closest in time.
Elevating large language model reasoning ability with auto-enhanced zero-shot prompts
Yuzhou Tang, Yibing Zhan, Changtong Zan, Long Lan, and Yonggang Che · 2025
Closest in time.
Optimizing generative ai by backpropagating language model feedback
Mert Yuksekgonul, Federico Bianchi, Joseph Boen, Sheng Liu, Pan Lu, Zhi Huang, Carlos Guestrin, and James Zou · 2025
Closest in time.
Toolrl: Reward is all tool learning needs
Cheng Qian, Emre Can Acikgoz, Qi He, Hongru Wang, Xiusi Chen, Dilek Hakkani-Tür, Gokhan Tur, and Heng Ji · 2025
Closest in time.
Self-evolved preference optimization for enhancing mathematical reasoning in small language models
Joykirat Singh, Tanmoy Chakraborty, and Akshay Nambi · 2025
Closest in time.
Meta-reasoning improves tool use in large language models
Lisa Alazraki and Marek Rei · 2025
Closest in time.
Multilingual llms inherently reward in-language time-sensitive semantic alignment for low-resource languages
Ashutosh Bajpai and Tanmoy Chakraborty · 2025
Closest in time.
Masrouter: Learning to route llms for multi-agent systems
Yanwei Yue, Guibin Zhang, Boyang Liu, Guancheng Wan, Kun Wang, Dawei Cheng, and Yiyan Qi · 2025
Closest in time.
The hidden dimensions of llm alignment: A multi-dimensional safety analysis
Wenbo Pan, Zhichao Liu, Qiguang Chen, Xiangyang Zhou, Haining Yu, and Xiaohua Jia · 2025
Closest in time.
A survey on multilingual large language models: Corpora, alignment, and bias
Yuemei Xu, Ling Hu, Jiayi Zhao, Zihan Qiu, Kexin Xu, Yuqi Ye, and Hanwen Gu · 2025
Closest in time.
Inference-time scaling for complex tasks: Where we stand and what lies ahead
Vidhisha Balachandran, Jingya Chen, Lingjiao Chen, Shivam Garg, Neel Joshi, Yash Lara, John Langford, Besmira Nushi, Vibhav Vineet, Yue Wu, et al · 2025
Closest in time.
Vapo: Efficient and reliable reinforcement learning for advanced reasoning tasks
Yufeng Yuan, Qiying Yu, Xiaochen Zuo, Ruofei Zhu, Wenyuan Xu, Jiaze Chen, Chengyi Wang, TianTian Fan, Zhengyin Du, Xiangpeng Wei, et al · 2025
Closest in time.
Seed-thinking-v1. 5: Advancing superb reasoning models with reinforcement learning
ByteDance Seed, Yufeng Yuan, Yu Yue, Mingxuan Wang, Xiaochen Zuo, Jiaze Chen, Lin Yan, Wenyuan Xu, Chi Zhang, Xin Liu, et al · 2025
Closest in time.
Efficient process reward model training via active learning
Keyu Duan, Zichen Liu, Xin Mao, Tianyu Pang, Changyu Chen, Qiguang Chen, Michael Qizhe Shieh, and Longxu Dou · 2025
Closest in time.
Stop overthinking: A survey on efficient reasoning for large language models
Yang Sui, Yu-Neng Chuang, Guanchu Wang, Jiamu Zhang, Tianyi Zhang, Jiayi Yuan, Hongyi Liu, Andrew Wen, Shaochen Zhong, Hanjie Chen, et al · 2025
Closest in time.
Thinkprune: Pruning long chain-of-thought of llms via reinforcement learning
Bairu Hou, Yang Zhang, Jiabao Ji, Yujian Liu, Kaizhi Qian, Jacob Andreas, and Shiyu Chang · 2025
Closest in time.