DELINE8K: A Synthetic Data Pipeline for the Semantic Segmentation of Historical Documents
Document semantic segmentation is a promising avenue that can facilitate document analysis tasks, including optical character recognition (OCR), form classification, and document editing. Although several synthetic datasets have been developed to distinguish handwriting from printed text, they fall short in class variety and document diversity. We demonstrate the limitations of training on existing datasets when solving the National Archives Form Semantic Segmentation dataset (NAFSS), a dataset which we introduce. To address these limitations, we propose the most comprehensive document semantic segmentation synthesis pipeline to date, incorporating preprinted text, handwriting, and document backgrounds from over 10 sources to create the Document Element Layer INtegration Ensemble 8K, or DELINE8K dataset. Our customized dataset exhibits superior performance on the NAFSS benchmark, demonstrating it as a promising tool in further research. The DELINE8K dataset is available at https://github.com/Tahlor/deline8k.
Code (0)
등록된 구현이 없습니다.
Tasks
8kDiversityFormOptical Character RecognitionOptical Character Recognition (OCR)SegmentationSemantic SegmentationSimilar Papers 제목 키워드 기반
A Data Augmentation Pipeline to Generate Synthetic Labeled Datasets of 3D Echocardiography Images using a GAN
Due to privacy issues and limited amount of publicly available labeled datasets in the domain of medical imaging, we propose an image generation pipeline to synthesize 3D echocardiographic images with corresponding groun…
Computed Tomography (CT)Data AugmentationGenerative Adversarial NetworkImage GenerationEnhanced Generative Data Augmentation for Semantic Segmentation via Stronger Guidance
Data augmentation is crucial for pixel-wise annotation tasks like semantic segmentation, where labeling requires significant effort and intensive labor. Traditional methods, involving simple transformations such as rotat…
Data AugmentationSegmentationSemantic SegmentationMulti-Contrast MRI Segmentation Trained on Synthetic Images
In our comprehensive experiments and evaluations, we show that it is possible to generate multiple contrast (even all synthetically) and use synthetically generated images to train an image segmentation engine. We showed…
Image SegmentationMRI segmentationSegmentationSemantic SegmentationVolume-based Semantic Labeling with Signed Distance Functions
Research works on the two topics of Semantic Segmentation and SLAM (Simultaneous Localization and Mapping) have been following separate tracks. Here, we link them quite tightly by delineating a category label fusion tech…
SegmentationSemantic SegmentationSimultaneous Localization and MappingSingle-Stage Semantic Segmentation from Image Labels
Recent years have seen a rapid growth in new approaches improving the accuracy of semantic segmentation in a weakly supervised setting, i.e. with only image-level labels available for training. However, this has come at …
SegmentationSemantic Segmentation