paper-with-me

Papers

Overview of FinMMEval 2026 Task 1: Multilingual Financial Multiple-Choice Question Answering

2026-07-22 · Zhuohan Xie, Yuyang Dai, Rania Elbadry, Vanshikaa Jani, Georgi Georgiev, Dimitar Dimitrov, Fan Zhang, Xueqing Peng, Lingfei Qian, Jimin Huang, Jiahui Geng, Yankai Chen, Ye Yuan, Haolun Wu, Yuxia Wang, Ivan Koychev, Veselin Stoyanov, Mingzi Song, Yu Chen, Xue Liu, Preslav Nakov arxiv

FinMMEval 2026 Task 1 evaluates multilingual financial multiple-choice question answering in English, Chinese, Arabic, and Hindi. The task tests whether systems can select the correct answer to finance questions involving domain terminology, numerical interpretation, and conceptual financial reasoning across languages and scripts. The final-test set contains 800 questions, with 200 questions per language; gold answers were withheld during submission, and each language was ranked independently by accuracy. The final leaderboards contain 13 English, 11 Chinese, 11 Arabic, and 10 Hindi ranked submissions. Top accuracies range from 92.0% in Hindi to 97.5% in English and Arabic, with the same leading teams appearing near the top across all four languages. The documented systems used retrieval augmentation, direct answer-option scoring, language-specific prompting, selective self-consistency, confidence checks, and LLM-based review stages.

📄 PDF Abstract BibTeX arXiv:2607.19856

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Overview of FinMMEval 2026 Task 2: Multilingual Financial Short-Answer Question Answering

2026-07-22 · Zhuohan Xie, Xueqing Peng, Georgi Georgiev, Dimitar Dimitrov 외 arxiv

FinMMEval 2026 Task 2 evaluates short-answer financial question answering over multilingual evidence. Each final-test item pairs an English question with financial statements and news in English, Chinese, Japanese, Spani…

Question Answering

The CLEF-2026 FinMMEval Lab: Multilingual and Multimodal Evaluation of Financial AI Systems

2026-02-11 · Zhuohan Xie, Rania Elbadry, Fan Zhang, Georgi Georgiev 외 arxiv

We present the setup and the tasks of the FinMMEval Lab at CLEF 2026, which introduces the first multilingual and multimodal evaluation framework for financial Large Language Models (LLMs). While recent advances in finan…

Question AnsweringDecision Making

Revolutionizing Finance with LLMs: An Overview of Applications and Insights

2024-01-22 · Huaqin Zhao, Zhengliang Liu, Zihao Wu, Yiwei Li 외

In recent years, Large Language Models (LLMs) like ChatGPT have seen considerable advancements and have been applied in diverse fields. Built on the Transformer architecture, these models are trained on extensive dataset…

A Survey of Large Language Models in Finance (FinLLMs)

2024-02-04 · Jean Lee, Nicholas Stevens, Soyeon Caren Han, Minseok Song

Large Language Models (LLMs) have shown remarkable capabilities across a wide variety of Natural Language Processing (NLP) tasks and have attracted attention from multiple domains, including financial services. Despite t…

Named Entity Recognition (NER)Question AnsweringSentiment ClassificationSurvey+3

MultiFinBen: A Multilingual, Multimodal, and Difficulty-Aware Benchmark for Financial LLM Evaluation

2025-06-16 · Xueqing Peng, Lingfei Qian, Yan Wang, Ruoyu Xiang 외

Recent advances in large language models (LLMs) have accelerated progress in financial NLP and applications, yet existing benchmarks remain limited to monolingual and unimodal settings, often over-relying on simple tasks…

Optical Character Recognition (OCR)