Fetching the paper…
Reading the bibliography…
In recent years, general-purpose large language models (LLMs) such as GPT, Gemini, Claude, and DeepSeek have advanced at an unprecedented pace.
FinQA: A dataset of numerical reasoning over financial data
Zhiyu Chen, Wenhu Chen, Charese Smiley, Sameena Shah, Iana Borova, Dylan Langdon, Reema Moussa, Matt Beane, Ting-Hao Huang, Bryan R. Routledge, and William Yang Wang · 2021
Earlier work this paper cites.
ConvFinQA: Exploring the chain of numerical reasoning in conversational finance question answering
Zhiyu Chen, Shiyang Li, Charese Smiley, Zhiqiang Ma, Sameena Shah, and William Yang Wang · 2022
Earlier work this paper cites.
Dakuan Lu, Hengkui Wu, Jiaqing Liang, Yipei Xu, Qianyu He, Yipeng Geng, Mengkun Han, Yingsi Xin, and Yanghua Xiao · 2023
Earlier work this paper cites.
Alpha-gpt: Human-ai interactive alpha mining for quantitative investment
Saizhuo Wang, Hang Yuan, Leon Zhou, Lionel M Ni, Heung-Yeung Shum, and Jian Guo · 2023
Earlier work this paper cites.
MMBench: Benchmarking end-to-end multimodal dnns and understanding their hardware-software implications
Cheng Xu, Xiaofeng Hou, Jiacheng Liu, Chao Li, Tianhao Huang, and Xiaozhi et al. Zhu · 2023
Earlier work this paper cites.
Universalner: Targeted distillation from large language models for open named entity recognition
Wenxuan Zhou, Sheng Zhang, Yu Gu, Muhao Chen, and Hoifung Poon · 2023
Earlier work this paper cites.
Financial Evaluation Dataset, 2023
Alipay Team · 2024
Earlier work this paper cites.
Twitter Financial News Sentiment, 2024
Anonymous · 2024
Earlier work this paper cites.
Fnspid: A comprehensive financial news dataset in time series
Zihan Dong, Xinyu Fan, and Zhiyuan Peng · 2024
Earlier work this paper cites.
FinCorpus, 2023a
Duxiaoman DI Team · 2024
Earlier work this paper cites.
FinanceIQ, 2023b
Duxiaoman DI Team · 2024
Earlier work this paper cites.
XuanYuan-finx1-preview, 2024
Duxiaoman DI Team · 2024
Cited alongside, same era.
Can large language models beat wall street? unveiling the potential of ai in stock selection
Georgios Fatouros, Konstantinos Metaxas, John Soldatos, and Dimosthenis Kyriazis · 2024
Cited alongside, same era.
Fineval: A chinese financial domain knowledge evaluation benchmark for large language models
Xin Guo, Haotian Xia, Zhaowei Liu, Hanyang Cao, Zhi Yang, Zhiqiang Liu, Sizhe Wang, Jinyi Niu, Chuqi Wang, Yanhui Wang, Xiaolong Liang, Xiaoming Huang, Bing Zhu, Zhongyu Wei, Yun Chen, Weining Shen, and Liwen Zhang · 2024
Cited alongside, same era.
Alphafin: Benchmarking financial analysis with retrieval-augmented stock-chain framework
Xiang Li, Zhenyu Li, Chen Shi, Yong Xu, Qing Du, Mingkui Tan, Jun Huang, and Wei Lin · 2024
Cited alongside, same era.
Quant-trading-instruct, 2024
Lukas Malik · 2024
Famma: A benchmark for financial domain multilingual multimodal question answering
Siqiao Xue, Tingting Chen, Fan Zhou, Qingyang Dai, Zhixuan Chu, and Hongyuan Mei · 2024
Later among the works it cites.
An Yang, Baosong Yang, Beichen Zhang, Binyuan Hui, Bo Zheng, Bowen Yu, Chengyuan Li, Dayiheng Liu, Fei Huang, Haoran Wei, et al · 2024
Later among the works it cites.
FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Design
Yangyang Yu, Haohang Li, Zhi Chen, Yuechen Jiang, Yang Li, and Denghui et al. Zhang · 2024
Later among the works it cites.
Finagent: A multimodal foundation agent for financial trading: Tool-augmented, diversified, and generalist
Wentao Zhang, Lingxuan Zhao, Haochong Xia, Shuo Sun, Jiaze Sun, Molei Qin, Xinyi Li, Yuqing Zhao, Yilei Zhao, Xinyu Cai, et al · 2024
Later among the works it cites.
JudgeLM : Fine-tuned large language models are scalable judges, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Qwq: Reflect deeply on the boundaries of the unknown., 2024
Qwen · 2024
Cited alongside, same era.
Qwen Team · 2024
Cited alongside, same era.
DeepSeekMath: Pushing the limits of mathematical reasoning in open language models
Zhihong Shao, Peiyi Wang, Qihao Zhu, Runxin Xu, Junxiao Song, Xiao Bi, and Haowei Zhang et al · 2024
Cited alongside, same era.
Ploutos: Towards interpretable stock movement prediction with financial large language model
Hanshuang Tong, Jun Li, Ning Wu, Ming Gong, Dongmei Zhang, and Qi Zhang · 2024
Cited alongside, same era.
Finnlp-agentscen-2024 shared task: Financial challenges in large language models-finllms
Qianqian Xie, Jimin Huang, Dong Li, Zhengyu Chen, Ruoyu Xiang, Mengxi Xiao, Yangyang Yu, Vijayasai Somasundaram, Kailai Yang, Chenhan Yuan, et al · 2024
Cited alongside, same era.
Learning to reason with llms., 2024a
OpenAI
Cited in the paper.
Mineru: An open-source solution for precise document content extraction, 2024a
Bin Wang, Chao Xu, Xiaomeng Zhao, Linke Ouyang, Fan Wu, Zhiyuan Zhao, Rui Xu, Kaiwen Liu, Yuan Qu, Fukai Shang, Bo Zhang, Liqun Wei, Zhihao Sui, Wei Li, Botian Shi, Yu Qiao, Dahua Lin, and Conghui He
Cited in the paper.
Lianghui Zhu, Xinggang Wang, and Xinlong Wang · 2024
Later among the works it cites.
Finance Instruct 500k, 2025
Joseph G. Flowers · 2025
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al · 2025
Closest in time.
Gpt-4o model documentation: Parameter configuration and data formats
OpenAI · 2025
Closest in time.
Fino1: On the transferability of reasoning enhanced llms to finance
Lingfei Qian, Weipeng Zhou, Yan Wang, Xueqing Peng, Jimin Huang, and Qianqian Xie · 2025
Closest in time.
Finarena: A human-agent collaboration framework for financial market analysis and forecasting
Congluo Xu, Zhaobin Liu, and Ziyang Li · 2025
Closest in time.