paper-with-me

홈 › Papers

A Single-Shot Arbitrarily-Shaped Text Detector based on Context Attended Multi-Task Learning

2019-08-15 · Pengfei Wang, Chengquan Zhang, Fei Qi, Zuming Huang, Mengyi En, Junyu Han, Jingtuo Liu, Errui Ding, Guangming Shi

Detecting scene text of arbitrary shapes has been a challenging task over the past years. In this paper, we propose a novel segmentation-based text detector, namely SAST, which employs a context attended multi-task learning framework based on a Fully Convolutional Network (FCN) to learn various geometric properties for the reconstruction of polygonal representation of text regions. Taking sequential characteristics of text into consideration, a Context Attention Block is introduced to capture long-range dependencies of pixel information to obtain a more reliable segmentation. In post-processing, a Point-to-Quad assignment method is proposed to cluster pixels into text instances by integrating both high-level object knowledge and low-level pixel information in a single shot. Moreover, the polygonal representation of arbitrarily-shaped text can be extracted with the proposed geometric properties much more effectively. Experiments on several benchmarks, including ICDAR2015, ICDAR2017-MLT, SCUT-CTW1500, and Total-Text, demonstrate that SAST achieves better or comparable performance in terms of accuracy. Furthermore, the proposed algorithm runs at 27.63 FPS on SCUT-CTW1500 with a Hmean of 81.0% on a single NVIDIA Titan Xp graphics card, surpassing most of the existing segmentation-based methods.

📄 PDF Abstract BibTeX arXiv:1908.05498

Code (1)

PaddlePaddle/PaddleOCR paddle

Tasks

Multi-Task LearningOptical Character Recognition (OCR)Scene Text DetectionSegmentation

Similar Papers 제목 키워드 기반

DEER: Detection-agnostic End-to-End Recognizer for Scene Text Spotting

2022-03-10 · Seonghyeon Kim, Seung Shin, Yoonsik Kim, Han-Cheol Cho 외

Recent end-to-end scene text spotters have achieved great improvement in recognizing arbitrary-shaped text instances. Common approaches for text spotting use region of interest pooling or segmentation masks to restrict f…

DecoderText Spotting

PGNet: Real-time Arbitrarily-Shaped Text Spotting with Point Gathering Network

2021-04-12 · Pengfei Wang, Chengquan Zhang, Fei Qi, Shanshan Liu 외

The reading of arbitrarily-shaped text has received increasing research attention. However, existing text spotters are mostly built on two-stage frameworks or character-based methods, which suffer from either Non-Maximum…

DecoderOptical Character Recognition (OCR)Scene Text DetectionText Spotting

FAST: Faster Arbitrarily-Shaped Text Detector with Minimalist Kernel Representation

2021-11-03 · Zhe Chen, Jiahao Wang, Wenhai Wang, Guo Chen 외

We propose an accurate and efficient scene text detection framework, termed FAST (i.e., faster arbitrarily-shaped text detector). Different from recent advanced text detectors that used complicated post-processing and ha…

GPUimage-classificationImage ClassificationScene Text Detection+1

SwinTextSpotter: Scene Text Spotting via Better Synergy between Text Detection and Text Recognition

2022-03-19 · CVPR 2022 1 · Mingxin Huang, Yuliang Liu, Zhenghao Peng, Chongyu Liu 외

End-to-end scene text spotting has attracted great attention in recent years due to the success of excavating the intrinsic synergy of the scene text detection and recognition. However, recent state-of-the-art methods us…

Scene Text DetectionText DetectionText Spotting

TransforMARS: Fault-Tolerant Self-Reconfiguration for Arbitrarily Shaped Modular Aerial Robot Systems

2025-09-17 · Rui Huang, Zhiyu Gao, Siyu Tang, Jialin Zhang 외 arxiv

Modular Aerial Robot Systems (MARS) consist of multiple drone modules that are physically bound together to form a single structure for flight. Exploiting structural redundancy, MARS can be reconfigured into different fo…