AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Regular Paper

CroFinBen: A Multilingual Benchmark of Large Language Models Across High- and Low-Resource Languages in Finance

School of Information Science and Engineering, Yunnan University, Kunming 650091, China
School of Artificial Intelligence, Wuhan University, Wuhan 430072, China
Show Author Information

Abstract

Recent advancements in financial large language models (FinLLMs) have shown strong performance in high-resource languages like English and Chinese, but are limited in low-resource settings, particularly in Southeast Asia (SEA), where labeled data resources are extremely scarce and challenging to annotate. Existing benchmarks primarily focus on high-resource finance, neglecting low-resource finance. To address these issues, we introduce CroFinBen, to the best of our knowledge, the first multilingual benchmark specifically designed to bridge the language-resource gap between high- and low-resource finance. It includes four key financial NLP tasks—financial sentiment analysis (FinSA), financial stock prediction (FinSP), financial text summarization (FinTS), and financial text classification (FinTC)—across both high-resource languages (English and Chinese) and low-resource Southeast Asian (SEA) languages (Indonesian, Malay, Thai, Filipino, and Vietnamese), comprising over 50 000 samples from 16 datasets, providing a comprehensive and balanced evaluation of LLMs. Unlike others that rely on full translations or overlook local context, CroFinBen incorporates localized annotations to reflect financial terms and cultural nuances in SEA languages. Evaluating 25 LLMs shows significant performance differences, with no clear proficiency in either high- or low-resource languages, especially for existing language-biased fine-tuned FinLLMs. The 1800B large-parameter closed-source GPT-4o excels, while DeepSeek-V3 and ChatGPT-3.5 also perform well. By bridging language-resource barriers, CroFinBen enhances the fairness and robustness of FinLLMs, providing strong support for improving performance in global financial scenarios. Our data resources are available at https://jcst.ict.ac.cn/en/supplement/afc97c89-905b-4397-941c-60dd0c720248.

Electronic Supplementary Material

Download File(s)
JCST-2505-15524-Highlights.pdf (319.3 KB)

References

【1】
【1】
 
 
Journal of Computer Science and Technology
Pages 977-992

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Hu G, Lyu S-Q, Wang Q-Q, et al. CroFinBen: A Multilingual Benchmark of Large Language Models Across High- and Low-Resource Languages in Finance. Journal of Computer Science and Technology, 2026, 41(3): 977-992. https://doi.org/10.1007/s11390-025-5524-7

4

Views

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 03 May 2025
Accepted: 07 November 2025
Published: 01 May 2026
© Institute of Computing Technology, Chinese Academy of Sciences 2026