Beyond Classification: Financial Reasoning in State-of-the-Art Language Models

Son, Guijin; Jung, Hanearl; Hahm, Moonjeong; Na, Keonju; Jin, Sol

Computer Science > Computation and Language

arXiv:2305.01505 (cs)

[Submitted on 30 Apr 2023 (v1), last revised 25 Jun 2023 (this version, v2)]

Title:Beyond Classification: Financial Reasoning in State-of-the-Art Language Models

Authors:Guijin Son, Hanearl Jung, Moonjeong Hahm, Keonju Na, Sol Jin

View PDF

Abstract:Large Language Models (LLMs), consisting of 100 billion or more parameters, have demonstrated remarkable ability in complex multi-step reasoning tasks. However, the application of such generic advancements has been limited to a few fields, such as clinical or legal, with the field of financial reasoning remaining largely unexplored. To the best of our knowledge, the ability of LLMs to solve financial reasoning problems has never been dealt with, and whether it can be performed at any scale remains unknown. To address this knowledge gap, this research presents a comprehensive investigation into the potential application of LLMs in the financial domain. The investigation includes a detailed exploration of a range of subjects, including task formulation, synthetic data generation, prompting methods, and evaluation capability. Furthermore, the study benchmarks various GPT variants with parameter scales ranging from 2.8B to 13B, with and without instruction tuning, on diverse dataset sizes. By analyzing the results, we reveal that the ability to generate coherent financial reasoning first emerges at 6B parameters, and continues to improve with better instruction-tuning or larger datasets. Additionally, the study provides a publicly accessible dataset named sFIOG (Synthetic-Financial Investment Opinion Generation), consisting of 11,802 synthetic investment thesis samples, to support further research in the field of financial reasoning. Overall, this research seeks to contribute to the understanding of the efficacy of language models in the field of finance, with a particular emphasis on their ability to engage in sophisticated reasoning and analysis within the context of investment decision-making.

Comments:	Accepted by FinNLP (Financial Technology and Natural Language Processing) @ IJCAI2023 as long paper
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY)
Cite as:	arXiv:2305.01505 [cs.CL]
	(or arXiv:2305.01505v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2305.01505

Submission history

From: Guijin Son [view email]
[v1] Sun, 30 Apr 2023 04:36:05 UTC (53 KB)
[v2] Sun, 25 Jun 2023 18:06:25 UTC (55 KB)

Computer Science > Computation and Language

Title:Beyond Classification: Financial Reasoning in State-of-the-Art Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Beyond Classification: Financial Reasoning in State-of-the-Art Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators