Fetching the paper…
Reading the bibliography…
LLMs have transformed NLP and shown promise in various fields, yet their potential in finance is underexplored due to a lack of comprehensive evaluation benchmarks, the rapid development of LLMs, and the complexity of financial tasks.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
FinBERT: Financial Sentiment Analysis with Pre-trained Language Models
Dogu Araci. 2019 · 1908
Earlier work this paper cites.
WWW’18 Open Challenge: Financial Opinion Mining and Question Answering
Macedo Maia, Siegfried Handschuh, Andre Freitas, Brian Davis, Ross McDermott, Manel Zarrouk, and Alexandra Balahur. 2018 · 1942
Earlier work this paper cites.
Stock movement prediction from tweets and historical prices. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 1970–1979
Yumo Xu and Shay B Cohen. 2018 · 1979
Earlier work this paper cites.
A monthly effect in stock returns
Robert A Ariel. 1987 · 1987
Earlier work this paper cites.
Statlog (German Credit Data)
Hans Hofmann. 1994 · 1994
Earlier work this paper cites.
Introduction to financial forecasting
Yaser S Abu-Mostafa and Amir F Atiya. 1996 · 1996
Earlier work this paper cites.
The sharpe ratio
William F Sharpe. 1998 · 1998
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries. In Text summarization branches out . 74–81
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Maximum drawdown
Malik Magdon-Ismail and Amir F Atiya. 2004 · 2004
Earlier work this paper cites.
A probabilistic interpretation of precision, recall and F-score, with implication for evaluation. In European conference on information retrieval . Springer, 345–359
Cyril Goutte and Eric Gaussier. 2005 · 2005
Earlier work this paper cites.
FinBERT: A Pretrained Language Model for Financial Communications
Yi Yang, Mark Christopher Siy UY, and Allen Huang. 2020b · 2006
Earlier work this paper cites.
Information extraction in finance . Vol. 8
Marco Costantino and Paolo Coletti. 2008 · 2008
Earlier work this paper cites.
Impact of News on the Commodity Market: Dataset and Results
Ankur Sinha and Tanmay Khandait. 2020 · 2009
Earlier work this paper cites.
Good debt or bad debt: Detecting semantic orientations in economic texts
Pekka Malo, Ankur Sinha, Pekka Korhonen, Jyrki Wallenius, and Pyry Takala. 2014 · 2014
Earlier work this paper cites.
Domain adaption of named entity recognition to support credit risk assessment. In Proceedings of the Australasian Language Technology Association Workshop 2015 . 84–90
Julio Cesar Salinas Alvarado, Karin Verspoor, and Timothy Baldwin. 2015 · 2015
Earlier work this paper cites.
Domain Adaption of Named Entity Recognition to Support Credit Risk Assessment. In Proceedings of the Australasian Language Technology Association Workshop 2015 , Ben Hachey and Kellie Webster (Eds.). Parramatta, Australia, 84–90
Julio Cesar Salinas Alvarado, Karin Verspoor, and Timothy Baldwin. 2015 · 2015
Earlier work this paper cites.
Complementarity, F-score, and NLP Evaluation. In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC’16) , Nicoletta Calzolari, Khalid Choukri, Thierry Declerck, Sara Goggi, Marko Grobelnik, Bente Maegaard, Joseph Mariani, Helene Mazo, Asuncion Moreno, Jan Odijk, and Stelios Piperidis (Eds.). European Language Resources Association (ELRA), Portorož, Slovenia, 261–266
Leon Derczynski. 2016 · 2016
Earlier work this paper cites.
Semeval-2017 task 5: Fine-grained sentiment analysis on financial microblogs and news. In Proceedings of the 11th international workshop on semantic evaluation (SemEval-2017) . 519–535
Keith Cortis, André Freitas, Tobias Daudert, Manuela Huerlimann, Manel Zarrouk, Siegfried Handschuh, and Brian Davis. 2017 · 2017
Earlier work this paper cites.
Strategic management decision-making in a complex world: quantifying, understanding, and using trade-offs
André E Punt. 2017 · 2017
Earlier work this paper cites.
Textual analogy parsing: What’s shared and what’s compared among analogous facts
Matthew Lamm, Arun Tejasvi Chaganty, Christopher D Manning, Dan Jurafsky, and Percy Liang. 2018 · 2018
Earlier work this paper cites.
Hybrid deep sequential modeling for social text-driven stock prediction. In Proceedings of the 27th ACM international conference on information and knowledge management . 1627–1630
Huizhe Wu, Wei Zhang, Weiwei Shen, and Jun Wang. 2018 · 2018
Earlier work this paper cites.
Machine learning and AI for risk management
Saqib Aziz and Michael Dowling. 2019 · 2019
Earlier work this paper cites.
Decision-making for financial trading: A fusion approach of machine learning and portfolio selection
Felipe Dias Paiva, Rodrigo Tomás Nogueira Cardoso, Gustavo Peixoto Hanaoka, and Wendel Moreira Duarte. 2019 · 2019
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q Weinberger, and Yoav Artzi. 2019 · 2019
Earlier work this paper cites.
The advantages of the Matthews correlation coefficient (MCC) over F1 score and accuracy in binary classification evaluation
Davide Chicco and Giuseppe Jurman. 2020 · 2020
Earlier work this paper cites.
End-to-end training for financial report summarization. In Proceedings of the 1st Joint Workshop on Financial Narrative Processing and MultiLing Financial Summarisation . 118–123
Moreno La Quatra and Luca Cagliero. 2020 · 2020
Earlier work this paper cites.
FinBERT: A Pre-trained Financial Language Representation Model for Financial Text Mining. In Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, IJCAI-20 , Christian Bessiere (Ed.). International Joint Conferences on Artificial Intelligence Organization, 4513–4519
Zhuang Liu, Degen Huang, Kaiyu Huang, Zhuang Li, and Jun Zhao. 2020 · 2020
Cited alongside, same era.
Textual analysis in finance
Tim Loughran and Bill McDonald. 2020 · 2020
Cited alongside, same era.
Financial document causality detection shared task (fincausal 2020)
Dominique Mariko, Hanna Abi Akl, Estelle Labidurie, Stephane Durfort, Hugues De Mazancourt, and Mahmoud El-Haj. 2020 · 2020
Cited alongside, same era.
Linyi Yang, Eoin M Kenny, Tin Lok James Ng, Yi Yang, Barry Smyth, and Ruihai Dong. 2020a · 2020
Cited alongside, same era.
Bizbench: A quantitative reasoning benchmark for business and finance
Rik Koncel-Kedziorski, Michael Krumdick, Viet Lai, Varshini Reddy, Charles Lovering, and Chris Tanner. 2023 · 2023
Later among the works it cites.
CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model
Yang Lei, Jiangtong Li, Ming Jiang, Junjie Hu, Dawei Cheng, Zhijun Ding, and Changjun Jiang. 2023 · 2023
Later among the works it cites.
Xianzhi Li, Xiaodan Zhu, Zhiqiang Ma, Xiaomo Liu, and Sameena Shah. 2023c · 2023
Later among the works it cites.
Fingpt: Democratizing internet-scale data for financial large language models
Xiao-Yang Liu, Guoxuan Wang, and Daochen Zha. 2023a · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
FinQA: A Dataset of Numerical Reasoning over Financial Data. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . 3697–3711
Zhiyu Chen, Wenhu Chen, Charese Smiley, Sameena Shah, Iana Borova, Dylan Langdon, Reema Moussa, Matt Beane, Ting-Hao Huang, Bryan R Routledge, et al · 2021
Cited alongside, same era.
Impact of news on the commodity market: Dataset and results. In Advances in Information and Communication: Proceedings of the 2021 Future of Information and Communication Conference (FICC), Volume 2 . Springer, 589–601
Ankur Sinha and Tanmay Khandait. 2021 · 2021
Cited alongside, same era.
Bartscore: Evaluating generated text as text generation
Weizhe Yuan, Graham Neubig, and Pengfei Liu. 2021 · 2021
Cited alongside, same era.
Trade the Event: Corporate Events Detection for News-Based Event-Driven Trading
Zhihan Zhou, Liqian Ma, and Han Liu. 2021 · 2021
Cited alongside, same era.
TAT-QA: A question answering benchmark on a hybrid of tabular and textual content in finance
Fengbin Zhu, Wenqiang Lei, Youcheng Huang, Chao Wang, Shuo Zhang, Jiancheng Lv, Fuli Feng, and Tat-Seng Chua. 2021 · 2021
Cited alongside, same era.
Ai in finance: challenges, techniques, and opportunities
Longbing Cao. 2022 · 2022
Cited alongside, same era.
GLM: General Language Model Pretraining with Autoregressive Blank Infilling. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 320–335
Zhengxiao Du, Yujie Qian, Xiao Liu, Ming Ding, Jiezhong Qiu, Zhilin Yang, and Jie Tang. 2022 · 2022
Cited alongside, same era.
FinRL-Meta: Market Environments and Benchmarks for Data-Driven Financial Reinforcement Learning
Xiao-Yang Liu, Ziyi Xia, Jingyang Rui, Jiechao Gao, Hongyang Yang, Ming Zhu, Christina Dan Wang, Zhaoran Wang, and Jian Guo. 2022 · 2022
Cited alongside, same era.
Alejandro Lopez-Lira and Yuehua Tang. 2023 · 2023
Later among the works it cites.
Gpt-4 technical report. arxiv 2303.08774
R OpenAI. 2023b · 2023
Later among the works it cites.
Code llama: Open foundation models for code
Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al · 2023
Later among the works it cites.
Trillion Dollar Words: A New Financial Dataset, Task & Market Analysis. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , Anna Rogers, Jordan Boyd-Graber, and Naoaki Okazaki (Eds.). Association for Computational Linguistics, Toronto, Canada, 6664–6679
Agam Shah, Suvan Paturi, and Sudheer Chava. 2023a · 2023
Later among the works it cites.
Finer: Financial named entity recognition dataset and weak-supervision model
Agam Shah, Ruchit Vithani, Abhinav Gullapalli, and Sudheer Chava. 2023b · 2023
Later among the works it cites.
Financial Numeric Extreme Labelling: A dataset and benchmarking. In Findings of the Association for Computational Linguistics: ACL 2023 . 3550–3561
Soumya Sharma, Subhendu Khatuya, Manjunath Hegde, Afreen Shaikh, Koustuv Dasgupta, Pawan Goyal, and Niloy Ganguly. 2023 · 2023
Later among the works it cites.
Fine-Grained Argument Understanding with BERT Ensemble Techniques: A Deep Dive into Financial Sentiment Analysis. In Proceedings of the 35th Conference on Computational Linguistics and Speech Processing (ROCLING 2023) . 242–249
Eugene Sy, Tzu-Cheng Peng, Shih-Hsuan Huang, Heng-Yu Lin, and Yung-Chun Chang. 2023 · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Later among the works it cites.
Internlm: A multilingual language model with progressively enhanced capabilities
InternLM Team. 2023 · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Later among the works it cites.
FinGPT: Instruction Tuning Benchmark for Open-Source Large Language Models in Financial Datasets
Neng Wang, Hongyang Yang, and Christina Dan Wang. 2023 · 2023
Later among the works it cites.
BloombergGPT: A Large Language Model for Finance
Shijie Wu, Ozan Irsoy, Steven Lu, Vadim Dabravolski, Mark Dredze, Sebastian Gehrmann, Prabhanjan Kambadur, David Rosenberg, and Gideon Mann. 2023 · 2023
Later among the works it cites.
Qianqian Xie, Weiguang Han, Yanzhao Lai, Min Peng, and Jimin Huang. 2023a · 2023
Later among the works it cites.
PIXIU: A Large Language Model, Instruction Data and Evaluation Benchmark for Finance
Qianqian Xie, Weiguang Han, Xiao Zhang, Yanzhao Lai, Min Peng, Alejandro Lopez-Lira, and Jimin Huang. 2023b · 2023
Later among the works it cites.
FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Design
Yangyang Yu, Haohang Li, Zhi Chen, Yuechen Jiang, Yang Li, Denghui Zhang, Rong Liu, Jordan W. Suchow, and Khaldoun Khashanah. 2023 · 2023
Later among the works it cites.
Forecasting the equity premium: Do deep neural network models work?
Xianzheng Zhou, Hui Zhou, and Huaigang Long. 2023 · 2023
Later among the works it cites.
LAiW: A Chinese Legal Large Language Models Benchmark
Yongfu Dai, Duanyu Feng, Jimin Huang, Haochen Jia, Qianqian Xie, Yifang Zhang, Weiguang Han, Wei Tian, and Hao Wang. 2024 · 2024
Closest in time.
Albert Q Jiang, Alexandre Sablayrolles, Antoine Roux, Arthur Mensch, Blanche Savary, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Emma Bou Hanna, Florian Bressand, et al · 2024
Closest in time.
A Survey of Large Language Models in Finance (FinLLMs)
Jean Lee, Nicholas Stevens, Soyeon Caren Han, and Minseok Song. 2024 · 2024
Closest in time.
Dólares or Dollars? Unraveling the Bilingual Prowess of Financial LLMs Between Spanish and English
Xiao Zhang, Ruoyu Xiang, Chenhan Yuan, Duanyu Feng, Weiguang Han, Alejandro Lopez-Lira, Xiao-Yang Liu, Sophia Ananiadou, Min Peng, Jimin Huang, and Qianqian Xie. 2024 · 2024
Closest in time.
Revolutionizing finance with llms: An overview of applications and insights
Huaqin Zhao, Zhengliang Liu, Zihao Wu, Yiwei Li, Tianze Yang, Peng Shu, Shaochen Xu, Haixing Dai, Lin Zhao, Gengchen Mai, et al · 2024
Closest in time.