paper-with-me

홈 › Papers

PubTables-v2: A new large-scale dataset for full-page and multi-page table extraction

2025-12-11 · Brandon Smock, Valerie Faucon-Morin, Max Sokolov, Libin Liang, Tayyibah Khanam, Amrit Ramesh, Maury Courtland arxiv

Table extraction (TE) is a key challenge in document understanding. Traditional approaches detect tables first, then recognize their structure. Recently, interest has surged in developing methods, such as vision-language models (VLMs), to extract tables directly in their full page or document context. However, a lack of annotated data has made progress difficult to demonstrate. To address this, we create a new large-scale dataset, PubTables-v2. PubTables-v2 unifies TE across various levels of surrounding context and, notably, is the first benchmark for multi-page TE. Our evaluations reveal that while current frontier models strongly outperform ($+0.354\ \textrm{GriTS}_\textrm{Con}$) small models on the most complex task (full-document multi-page TE), this gap can be closed or even reversed ($-0.056\ \textrm{GriTS}_\textrm{Con}$) on narrower tasks (cropped table extraction) with targeted training. Data is available at https://huggingface.co/datasets/kensho/PubTables-v2. Code and models will be released.

📄 PDF Abstract BibTeX arXiv:2512.10888

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

POTATR: A Lightweight Image-to-Graph Model for Page-Level Table Extraction

2026-06-08 · Brandon Smock, Libin Liang, Max Sokolov, Amrit Ramesh 외 arxiv

Large-scale document processing requires contextually aware table extraction (TE) that is both accurate and efficient. Yet current approaches require billions of parameters, hundreds of autoregressive steps, or costly AP…

PubTables-1M: Towards comprehensive table extraction from unstructured documents

2021-09-30 · CVPR 2022 1 · Brandon Smock, Rohith Pesala, Robin Abraham

Recently, significant progress has been made applying machine learning to the problem of table structure inference and extraction from unstructured documents. However, one of the greatest challenges remains the creation …

Articlesobject-detectionObject DetectionTable Detection+3

A Hybrid Approach for Document Layout Analysis in Document images

2024-04-27 · Tahira Shehzadi, Didier Stricker, Muhammad Zeshan Afzal

Document layout analysis involves understanding the arrangement of elements within a document. This paper navigates the complexities of understanding various elements within document images, such as text, images, tables,…

Contrastive LearningDecoderDocument Layout AnalysisInformation Retrieval+3

CTE: A Dataset for Contextualized Table Extraction

2023-02-02 · Andrea Gemelli, Emanuele Vivoli, Simone Marinai

Relevant information in documents is often summarized in tables, helping the reader to identify useful facts. Most benchmark datasets support either document layout analysis or table understanding, but lack in providing …

Document Layout AnalysisTable DetectionTable Extraction

Aligning benchmark datasets for table structure recognition

2023-03-01 · Brandon Smock, Rohith Pesala, Robin Abraham

Benchmark datasets for table structure recognition (TSR) must be carefully processed to ensure they are annotated consistently. However, even if a dataset's annotations are self-consistent, there may be significant incon…

Table DetectionTable Recognition