paper-with-me

홈 › Papers

KoViDoRe: Korean Visual Document Retrieval

2026-08-21 · Yongbin Choi, Yongwoo Song, Mujeen Sung arxiv

Recent advances in multimodal retrieval have improved the ability to retrieve information from visually rich documents such as PDFs and reports. However, existing benchmarks remain largely centered on English and provide limited coverage of Korean visual documents with complex structures. Furthermore, most existing Korean resources primarily evaluate single-page retrieval, failing to capture realistic scenarios that require evidence aggregation across multiple pages. To address these gaps, we introduce KoViDoRe, a benchmark for Korean visual document retrieval. The dataset is constructed from publicly available Korean documents with diverse layouts, including tables, figures, and multi-column structures. We develop a multi-stage data curation pipeline consisting of structured document parsing, synthetic query generation using both summary-based and context-based strategies, and relevance mapping with human verification. Using KoViDoRe, we evaluate a wide range of multimodal retrieval models and observe that current models struggle to effectively handle Korean visual document retrieval, particularly in settings involving structured content and diverse query types. Motivated by this finding, we further curate a large-scale training dataset, Ko-VDR Train Public, to support the development of retrieval models tailored to Korean visual documents. Together, KoViDoRe and Ko-VDR Train Public provide a unified benchmark and training resource for Korean visual document retrieval.

📄 PDF Abstract BibTeX arXiv:2608.20840

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SDS KoPub VDR: A Benchmark Dataset for Visual Document Retrieval in Korean Public Documents

2025-11-07 · Jaehoon Lee, Sohyun Kim, Wanggeun Park, Geon Lee 외 arxiv

Existing benchmarks for visual document retrieval (VDR) largely overlook non-English languages and the structural complexity of official publications. To address this gap, we introduce SDS KoPub VDR, the first large-scal…

A Method of Passage-Based Document Retrieval in Question Answering System

2015-12-17 · Jong Man-Hung, Ri Chong-Han, Choe Hyok-Chol, Hwang Chol-Jun

We propose a method for using the scoring values of passages to effectively retrieve documents in a Question Answering system. For this, we suggest evaluation function that considers proximity between each question ter…

Question AnsweringRetrieval

HERITAGE: An End-to-End Web Platform for Processing Korean Historical Documents in Hanja

2025-01-21 · Seyoung Song, Haneul Yoo, Jiho Jin, Kyunghyun Cho 외

While Korean historical documents are invaluable cultural heritage, understanding those documents requires in-depth Hanja expertise. Hanja is an ancient language used in Korea before the 20th century, whose characters we…

document understandingMachine Translationnamed-entity-recognitionNamed Entity Recognition+1

Prompt-RAG: Pioneering Vector Embedding-Free Retrieval-Augmented Generation in Niche Domains, Exemplified by Korean Medicine

2024-01-20 · Bongsu Kang, Jundong Kim, Tae-Rim Yun, Chang-Eop Kim

We propose a natural language prompt-based retrieval augmented generation (Prompt-RAG), a novel approach to enhance the performance of generative large language models (LLMs) in niche domains. Conventional RAG methods mo…

ChatbotInformativenessQuestion AnsweringRAG+2

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

2026-05-28 · Eunbyeol Cho, Yunseung Lee, Mirae Kim, Jeewon Yang 외 arxiv

Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barrier to deployment in high-stakes environments. Existing benchmarks fo…