paper-with-me

Papers

Extraction of Virtual Baselines From Distorted Document Images Using Curvilinear Projection

2015-12-01 · ICCV 2015 12 · Gaofeng Meng, Zuming Huang, Yonghong Song, Shiming Xiang, Chunhong Pan

The baselines of a document page are a set of virtual horizontal and parallel lines, to which the printed contents of document, e.g., text lines, tables or inserted photos, are aligned. Accurate baseline extraction is of great importance in the geometric correction of curved document images. In this paper, we propose an efficient method for accurate extraction of these virtual visual cues from a curved document image. Our method comes from two basic observations that the baselines of documents do not intersect with each other and that within a narrow strip, the baselines can be well approximated by linear segments. Based upon these observations, we propose a curvilinear projection based method and model the estimation of curved baselines as a constrained sequential optimization problem. A dynamic programming algorithm is then developed to efficiently solve the problem. The proposed method can extract the complete baselines through each pixel of document images in a high accuracy. It is also scripts insensitive and highly robust to image noises, non-textual objects, image resolutions and image quality degradation like blurring and non-uniform illumination. Extensive experiments on a number of captured document images demonstrate the effectiveness of the proposed method.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Foreground and Text-lines Aware Document Image Rectification

2023-01-01 · ICCV 2023 1 · Heng Li, XiangPing Wu, Qingcai Chen, Qianjin Xiang

This paper aims at the distorted document image rectification problem, the objective to eliminate the geometric distortion in the document images and realize document intelligence. Improving the readability of distor…

Robustness of Structured Data Extraction from Perspectively Distorted Documents

2025-11-18 · Hyakka Nakada, Yoshiyasu Tanaka arxiv

Optical Character Recognition (OCR) for data extraction from documents is essential to intelligent informatics, such as digitizing medical records and recognizing road signs. Multi-modal Large Language Models (LLMs) can …

Deep Unrestricted Document Image Rectification

2023-04-18 · Hao Feng, Shaokai Liu, Jiajun Deng, Wengang Zhou 외

In recent years, tremendous efforts have been made on document image rectification, but existing advanced algorithms are limited to processing restricted document images, i.e., the input images must incorporate a complet…

Local Distortion

Dewarping Document Image By Displacement Flow Estimation with Fully Convolutional Network

2021-04-14 · Guo-Wang Xie, Fei Yin, Xu-Yao Zhang, Cheng-Lin Liu

As camera-based documents are increasingly used, the rectification of distorted document images becomes a need to improve the recognition performance. In this paper, we propose a novel framework for both rectifying disto…

Exploiting Vector Fields for Geometric Rectification of Distorted Document Images

2018-09-01 · ECCV 2018 9 · Gaofeng MENG, Yuanqi SU, Ying Wu, Shiming Xiang 외

This paper proposes a segment-free method for geometric rectification of a distorted document image captured by a hand-held camera. The method can recover the 3D page shape by exploiting the intrinsic vector fields of th…