paper-with-me

Papers

Extracting Tables from Documents using Conditional Generative Adversarial Networks and Genetic Algorithms

2019-04-03 · Nataliya Le Vine, Matthew Zeigenfuse, Mark Rowan

Extracting information from tables in documents presents a significant challenge in many industries and in academic research. Existing methods which take a bottom-up approach of integrating lines into cells and rows or columns neglect the available prior information relating to table structure. Our proposed method takes a top-down approach, first using a generative adversarial network to map a table image into a standardised `skeleton' table form denoting the approximate row and column borders without table content, then fitting renderings of candidate latent table structures to the skeleton structure using a distance measure optimised by a genetic algorithm.

📄 PDF Abstract BibTeX arXiv:1904.01947

Code (0)

등록된 구현이 없습니다.

Tasks

Generative Adversarial Network

Similar Papers 제목 키워드 기반

Identifying Table Structure in Documents using Conditional Generative Adversarial Networks

2020-01-13 · Nataliya Le Vine, Claus Horn, Matthew Zeigenfuse, Mark Rowan

In many industries, as well as in academic research, information is primarily transmitted in the form of unstructured documents (this article, for example). Hierarchically-related data is rendered as tables, and extracti…

Generative Adversarial NetworkSmall Data Image Classification

Row Conditional-TGAN for generating synthetic relational databases

2022-11-14 · Mohamed Gueye, Yazid Attabi, Maxime Dumas

Besides reproducing tabular data properties of standalone tables, synthetic relational databases also require modeling the relationships between related tables. In this paper, we propose the Row Conditional-Tabular Gener…

Generative Adversarial Network

A Conglomerate of Multiple OCR Table Detection and Extraction

2020-10-16 · Smita Pallavi, Raj Ratn Pranesh, Sumit Kumar

Information representation as tables are compact and concise method that eases searching, indexing, and storage requirements. Extracting and cloning tables from parsable documents is easier and widely used, however indus…

Optical Character Recognition (OCR)Table Detection

Extracting Body Text from Academic PDF Documents for Text Mining

2020-10-23 · Changfeng Yu, Cheng Zhang, Jie Wang

Accurate extraction of body text from PDF-formatted academic documents is essential in text-mining applications for deeper semantic understandings. The objective is to extract complete sentences in the body text into a t…

Sentence

Large-Scale Data Mining of Rapid Residue Detection Assay Data From HTML and PDF Documents: Improving Data Access and Visualization for Veterinarians

2021-12-02 · Majid Jaberi-Douraki, Soudabeh Taghian Dinani, Nuwan Indika Millagaha Gedara, Xuan Xu 외

Extra-label drug use in food animal medicine is authorized by the US Animal Medicinal Drug Use Clarification Act (AMDUCA), and estimated withdrawal intervals are based on published scientific pharmacokinetic data. Occasi…