Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated remarkable success in various tasks such as natural language understanding, text summarization, and machine translation.
Www’18 open challenge: financial opinion mining and question answering
Macedo Maia, Siegfried Handschuh, André Freitas, Brian Davis, Ross McDermott, Manel Zarrouk, and Alexandra Balahur. 2018 · 1942
Earlier work this paper cites.
A semantic representation for domain-specific patterns
Susana Montero, Paloma Díaz, and Ignacio Aedo. 2004 · 2004
Earlier work this paper cites.
Standards for health care decision-making: legal and practical considerations
Kim Dayton. 2012 · 2012
Earlier work this paper cites.
Bioasq: A challenge on large-scale biomedical semantic indexing and question answering
George Tsatsaronis and 1 others. 2012 · 2012
Earlier work this paper cites.
Good debt or bad debt: Detecting semantic orientations in economic texts
P. Malo, A. Sinha, P. Korhonen, J. Wallenius, and P. Takala. 2014 · 2014
Earlier work this paper cites.
Pubmed 200k rct: a dataset for sequential sentence classification in medical abstracts
Franck Dernoncourt and Ji Young Lee. 2017 · 2017
Earlier work this paper cites.
Relating chemistry to healthcare and more: implementation of more in a survey organic and biochemistry course for prehealth students
Lianne Schroeder, Joshua Bierdz, Donald J Wink, Maripat King, Patrick L Daubenmire, and Ginevra A Clark. 2018 · 2018
Earlier work this paper cites.
Pubmedqa: A dataset for biomedical research question answering
Qiao Jin, Bhuwan Dhingra, Zhengping Liu, William Cohen, and Xinghua Lu. 2019 · 2019
Earlier work this paper cites.
Latent retrieval for weakly supervised open domain question answering
Kenton Lee, Ming-Wei Chang, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
The financial narrative summarisation shared task (FNS 2020)
Mahmoud El-Haj, Ahmed AbuRa’ed, Marina Litvak, Nikiforos Pittaras, and George Giannakopoulos. 2020 · 2020
Earlier work this paper cites.
S2orc: The semantic scholar open research corpus
Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney, and Daniel S Weld. 2020 · 2020
Earlier work this paper cites.
On the effectiveness of adapter-based tuning for pretrained language model adaptation
Ruidan He, Linlin Liu, Hai Ye, Qingyu Tan, Bosheng Ding, Liying Cheng, Jiawei Low, Lidong Bing, and Luo Si. 2021 · 2021
Earlier work this paper cites.
What disease does this patient have? a large-scale open domain question answering dataset from medical exams
Di Jin, Eileen Pan, Nassim Oufattole, Wei-Hung Weng, Hanyi Fang, and Peter Szolovits. 2021 · 2021
Earlier work this paper cites.
Crossner: Evaluating cross-domain named entity recognition
Zihan Liu, Yan Xu, Tiezheng Yu, Wenliang Dai, Ziwei Ji, Samuel Cahyawijaya, Andrea Madotto, and Pascale Fung. 2021 · 2021
Earlier work this paper cites.
Psyqa: A chinese dataset for generating long counseling text for mental health support
Hao Sun, Zhenru Lin, Chujie Zheng, Siyang Liu, and Minlie Huang. 2021 · 2021
Earlier work this paper cites.
K-adapter: Infusing knowledge into pre-trained models with adapters
Ruize Wang, Duyu Tang, Nan Duan, Zhongyu Wei, Xuan-Jing Huang, Jianshu Ji, Guihong Cao, Daxin Jiang, and Ming Zhou. 2021 · 2021
Earlier work this paper cites.
Nlm at mediqa 2021: Transfer learning-based approaches for consumer question and multi-answer summarization
Shweta Yadav, Mourad Sarrouti, and Deepak Gupta. 2021 · 2021
Earlier work this paper cites.
Calibrating factual knowledge in pretrained language models
Qingxiu Dong, Damai Dai, Yifan Song, Jingjing Xu, Zhifang Sui, and Lei Li. 2022 · 2022
Earlier work this paper cites.
Meta-learning the difference: preparing large language models for efficient adaptation
Zejiang Hou, Julian Salazar, and George Polovets. 2022 · 2022
Earlier work this paper cites.
Lifelong pretraining: Continually adapting language models to emerging corpora
Xisen Jin, Dejiao Zhang, Henghui Zhu, Wei Xiao, Shang-Wen Li, Xiaokai Wei, Andrew Arnold, and Xiang Ren. 2022 · 2022
Earlier work this paper cites.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu. 2022 · 2022
Earlier work this paper cites.
Benchmarking multi-modal entailment for fact verification
Parth Patwa, Shreyash Mishra, S Suryavardan, Amrit Bhaskar, Parul Chopra, Aishwarya Reganti, Amitava Das, Tanmoy Chakraborty, Amit Sheth, Asif Ekbal, and 1 others. 2022 · 2022
Earlier work this paper cites.
When flue meets flang: Benchmarks and large pre-trained language model for financial domain
Raj Sanjay Shah and 1 others. 2022 · 2022
Earlier work this paper cites.
Multi-lexsum: Real-world summaries of civil rights lawsuits at multiple granularities
Zejiang Shen, Kyle Lo, Lauren Yu, Nathan Dahlberg, Margo Schlanger, and Doug Downey. 2022 · 2022
Earlier work this paper cites.
Aiseckg: Knowledge graph dataset for cybersecurity education
Garima Agrawal. 2023 · 2023
Earlier work this paper cites.
Chemcrow: Augmenting large-language models with chemistry tools
Andres M Bran, Sam Cox, Oliver Schilter, Carlo Baldassari, Andrew D White, and Philippe Schwaller. 2023 · 2023
Cited alongside, same era.
Soulchat: Improving llms’ empathy, listening, and comfort abilities through fine-tuning with multi-turn empathy conversations
Yirong Chen, Xiaofen Xing, Jingkai Lin, Huimin Zheng, Zhenyu Wang, Qi Liu, and Xiangmin Xu. 2023 · 2023
Cited alongside, same era.
Educhat: A large-scale language model-based chatbot system for intelligent education
Yuhao Dan, Zhikai Lei, Yiyang Gu, Yong Li, Jianghao Yin, Jiaju Lin, Linhao Ye, Zhiyan Tie, Yougen Zhou, Yilei Wang, and 1 others. 2023 · 2023
Cited alongside, same era.
Legalbench: A collaboratively built benchmark for measuring legal reasoning in large language models
Neel Guha, Julian Nyarko, Daniel Ho, Christopher Ré, Adam Chilton, Alex Chohlas-Wood, Austin Peters, Brandon Waldon, Daniel Rockmore, Diego Zambrano, and 1 others. 2023 · 2023
Cited alongside, same era.
A survey of knowledge enhanced pre-trained language models
Medinst: Meta dataset of biomedical instructions
Wenhan Han, Meng Fang, Zihan Zhang, Yu Yin, Zirui Song, Ling Chen, Mykola Pechenizkiy, and Qingyu Chen. 2024 · 2024
Later among the works it cites.
Adaptive-rag: Learning to adapt retrieval-augmented large language models through question complexity
Soyeong Jeong, Jinheon Baek, Sukmin Cho, Sung Ju Hwang, and Jong C Park. 2024 · 2024
Later among the works it cites.
Genegpt: Augmenting large language models with domain tools for improved access to biomedical information
Qiao Jin, Yifan Yang, Qingyu Chen, and Zhiyong Lu. 2024 · 2024
Later among the works it cites.
Finllama: Financial sentiment classification for algorithmic trading applications
Thanos Konstantinidis and 1 others. 2024 · 2024
Later among the works it cites.
Biot5+: Towards generalized biological understanding with iupac integration and multi-task tuning
Qizhi Pei and 1 others. 2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Linmei Hu, Zeyi Liu, Ziwang Zhao, Lei Hou, Liqiang Nie, and Juanzi Li. 2023 · 2023
Cited alongside, same era.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
Cited alongside, same era.
Social-llm: Modeling user behavior at scale using language models and social network data
Julie Jiang and Emilio Ferrara. 2023 · 2023
Cited alongside, same era.
Dakuan Lu, Hengkui Wu, Jiaqing Liang, Yipei Xu, Qianyu He, Yipeng Geng, Mengkun Han, Yingsi Xin, and Yanghua Xiao. 2023 · 2023
Cited alongside, same era.
Factscore: Fine-grained atomic evaluation of factual precision in long form text generation
Sewon Min, Kalpesh Krishna, Xinxi Lyu, Mike Lewis, Wen-tau Yih, Pang Koh, Mohit Iyyer, Luke Zettlemoyer, and Hannaneh Hajishirzi. 2023 · 2023
Cited alongside, same era.
Smile: Single-turn to multi-turn inclusive language expansion via chatgpt for mental health support
Huachuan Qiu, Hongliang He, Shuai Zhang, Anqi Li, and Zhenzhong Lan. 2023 · 2023
Cited alongside, same era.
Qiaoban: A parental emotion coaching dialogue assistant for better parent-child interaction
Zhao Weixiang, Wang Shilong, Tong Yanpeng, Lu Xin, Li Zhuojun, Zhao Yanyan, Wang Chenxue, Zheng Tian, and Qin Bing. 2023 · 2023
Cited alongside, same era.
Pixiu: a large language model, instruction data and evaluation benchmark for finance
Qianqian Xie, Weiguang Han, Xiao Zhang, Yanzhao Lai, Min Peng, Alejandro Lopez-Lira, and Jimin Huang. 2023 · 2023
Cited alongside, same era.
Tong Xie, Yuwei Wan, Yixuan Liu, Yuchen Zeng, Wenjie Zhang, Chunyu Kit, Dongzhan Zhou, and Bram Hoex. 2024 · 2024
Later among the works it cites.
Emollm: Multimodal emotional understanding meets large language models
Qu Yang, Mang Ye, and Bo Du. 2024 · 2024
Later among the works it cites.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan. 2024 · 2024
Later among the works it cites.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, He Ye, and Yanlin Wang. 2024 · 2024
Later among the works it cites.
K-comp: Retrieval-augmented medical domain question answering with knowledge-injected compressor
Jeonghun Cho and Gary Geunbae Lee. 2025 · 2025
Closest in time.
Llm alignment as retriever optimization: An information retrieval perspective
Bowen Jin, Jinsung Yoon, Zhen Qin, Ziqi Wang, Wei Xiong, Yu Meng, Jiawei Han, and Sercan O Arik. 2025 · 2025
Closest in time.
A fine-tuning approach for t5 using knowledge graphs to address complex tasks
Xiaoxuan Liao, Binrong Zhu, Jacky He, Guiran Liu, Hongye Zheng, and Jia Gao. 2025 · 2025
Closest in time.
Foundational large language models for materials research
Vaibhav Mishra, Somaditya Singh, Dhruv Ahlawat, Mohd Zaki, Vaibhav Bihani, Hargun Singh Grover, Biswajit Mishra, Santiago Miret, Mausam, and N. M. Anoop Krishnan. 2025 · 2025
Closest in time.
Soft prompt tuning for augmenting dense retrieval with large language models
Zhiyuan Peng, Xuyang Wu, Qifan Wang, and Yi Fang. 2025 · 2025
Closest in time.
Omniscience: A domain-specialized llm for scientific reasoning and discovery
Vignesh Prabhakar, Md Amirul Islam, Adam Atanas, Yao-Ting Wang, Joah Han, Aastha Jhunjhunwala, Rucha Apte, Robert Clark, Kang Xu, Zihan Wang, and Kai Liu. 2025 · 2025
Closest in time.
Fino1: On the transferability of reasoning enhanced llms to finance
Lingfei Qian, Weipeng Zhou, Yan Wang, Xueqing Peng, Han Yi, Jimin Huang, Qianqian Xie, and Jianyun Nie. 2025 · 2025
Closest in time.
Self-calibrated listwise reranking with large language models
Ruiyang Ren, Yuhao Wang, Kun Zhou, Wayne Xin Zhao, Wenjie Wang, Jing Liu, Ji-Rong Wen, and Tat-Seng Chua. 2025 · 2025
Closest in time.
Manish Sanwal. 2025 · 2025
Closest in time.
Toward expert-level medical question answering with large language models
Karan Singhal, Tao Tu, Juraj Gottweis, Rory Sayres, Ellery Wulczyn, Mohamed Amin, Le Hou, Kevin Clark, Stephen R Pfohl, Heather Cole-Lewis, and 1 others. 2025 · 2025
Closest in time.
Zirui Song, Guangxian Ouyang, Mingzhe Li, Yuheng Ji, Chenxi Wang, Zixiang Xu, Zeyu Zhang, Xiaoqing Zhang, Qian Jiang, Zhenhao Chen, and 1 others. 2025 · 2025
Closest in time.
Beyond profile: From surface-level facts to deep persona simulation in llms
Zixiao Wang, Duzhen Zhang, Ishita Agrawal, Shen Gao, Le Song, and Xiuying Chen. 2025 · 2025
Closest in time.
Socialmaze: A benchmark for evaluating social reasoning in large language models
Zixiang Xu, Yanbo Wang, Yue Huang, Jiayi Ye, Haomin Zhuang, Zirui Song, Lang Gao, Chenxi Wang, Zhaorun Chen, Yujun Zhou, and 1 others. 2025 · 2025
Closest in time.
Divide-fuse-conquer: Eliciting" aha moments" in multi-scenario games
Xiaoqing Zhang, Huabin Zheng, Ang Lv, Yuhan Liu, Zirui Song, Xiuying Chen, Rui Yan, and Flood Sung. 2025 · 2025
Closest in time.
Medrag: Enhancing retrieval-augmented generation with knowledge graph-elicited reasoning for healthcare copilot
Xuejiao Zhao, Siyan Liu, Su-Yin Yang, and Chunyan Miao. 2025 · 2025
Closest in time.