paper-with-me

홈 › Papers

Unveiling Document Structures with YOLOv5 Layout Detection

2023-09-29 · Herman Sugiharto, Yorissa Silviana, Yani Siti Nurpazrin

The current digital environment is characterized by the widespread presence of data, particularly unstructured data, which poses many issues in sectors including finance, healthcare, and education. Conventional techniques for data extraction encounter difficulties in dealing with the inherent variety and complexity of unstructured data, hence requiring the adoption of more efficient methodologies. This research investigates the utilization of YOLOv5, a cutting-edge computer vision model, for the purpose of rapidly identifying document layouts and extracting unstructured data. The present study establishes a conceptual framework for delineating the notion of "objects" as they pertain to documents, incorporating various elements such as paragraphs, tables, photos, and other constituent parts. The main objective is to create an autonomous system that can effectively recognize document layouts and extract unstructured data, hence improving the effectiveness of data extraction. In the conducted examination, the YOLOv5 model exhibits notable effectiveness in the task of document layout identification, attaining a high accuracy rate along with a precision value of 0.91, a recall value of 0.971, an F1-score of 0.939, and an area under the receiver operating characteristic curve (AUC-ROC) of 0.975. The remarkable performance of this system optimizes the process of extracting textual and tabular data from document images. Its prospective applications are not limited to document analysis but can encompass unstructured data from diverse sources, such as audio data. This study lays the foundation for future investigations into the wider applicability of YOLOv5 in managing various types of unstructured data, offering potential for novel applications across multiple domains.

📄 PDF Abstract BibTeX arXiv:2309.17033

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Latent Diffusion for Guided Document Table Generation

2024-08-19 · Syed Jawwad Haider Hamdani, Saifullah Saifullah, Stefan Agne, Andreas Dengel 외

Obtaining annotated table structure data for complex tables is a challenging task due to the inherent diversity and complexity of real-world document layouts. The scarcity of publicly available datasets with comprehensiv…

object-detectionObject Detection

Framework and Model Analysis on Bengali Document Layout Analysis Dataset: BaDLAD

2023-08-15 · Kazi Reyazul Hasan, Mubasshira Musarrat, Sadif Ahmed, Shahriar Raj

This study focuses on understanding Bengali Document Layouts using advanced computer programs: Detectron2, YOLOv8, and SAM. We looked at lots of different Bengali documents in our study. Detectron2 is great at finding an…

Document Layout Analysis

AutoFormBench: Benchmark Dataset for Automating Form Understanding

2026-03-31 · Gaurab Baral, Junxiu Zhou arxiv

Automated processing of structured documents such as government forms, healthcare records, and enterprise invoices remains a persistent challenge due to the high degree of layout variability encountered in real-world set…

Page Layout Analysis of Text-heavy Historical Documents: a Comparison of Textual and Visual Approaches

2022-12-12 · Najem-Meyer Sven, Romanello Matteo

Page layout analysis is a fundamental step in document processing which enables to segment a page into regions of interest. With highly complex layouts and mixed scripts, scholarly commentaries are text-heavy documents w…

Position

From Codicology to Code: A Comparative Study of Transformer and YOLO-based Detectors for Layout Analysis in Historical Documents

2025-06-25 · Sergio Torres Aguilar

Robust Document Layout Analysis (DLA) is critical for the automated processing and understanding of historical documents with complex page organizations. This paper benchmarks five state-of-the-art object detection archi…

Document Layout Analysisobject-detectionObject Detection