paper-with-me

Papers

Geometry-Aware Scene Text Detection With Instance Transformation Network

2018-06-01 · CVPR 2018 6 · Fangfang Wang, Liming Zhao, Xi Li, Xinchao Wang, DaCheng Tao

Localizing text in the wild is challenging in the situations of complicated geometric layout of the targets like random orientation and large aspect ratio. In this paper, we propose a geometry-aware modeling approach tailored for scene text representation with an end-to-end learning scheme. In our approach, a novel Instance Transformation Network (ITN) is presented to learn the geometry-aware representation encoding the unique geometric configurations of scene text instances with in-network transformation embedding, resulting in a robust and elegant framework to detect words or text lines at one pass. An end-to-end multi-task learning strategy with transformation regression, text/non-text classification and coordinate regression is adopted in the ITN. Experiments on the benchmark datasets demonstrate the effectiveness of the proposed approach in detecting scene text in various geometric configurations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

General ClassificationMulti-Task LearningregressionScene Text Detectiontext-classificationText ClassificationText Detection

Similar Papers 제목 키워드 기반

Geometry Normalization Networks for Accurate Scene Text Detection

2019-09-02 · ICCV 2019 10 · Youjiang Xu, Jiaqi Duan, Zhanghui Kuang, Xiaoyu Yue 외

Large geometry (e.g., orientation) variances are the key challenges in the scene text detection. In this work, we first conduct experiments to investigate the capacity of networks for learning geometry variances on detec…

Scene Text DetectionText Detection

Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection

2026-02-24 · Xiaokai Bai, Lianqing Zheng, Si-Yuan Cao, Xiaohan Zhang 외 arxiv

4D millimeter-wave radar has emerged as a promising sensing modality for autonomous driving due to its robustness and affordability. However, its sparse and weak geometric cues make reliable instance activation difficult…

Scene Understanding3D Object DetectionAutonomous Driving

DisARM: Displacement Aware Relation Module for 3D Detection

2022-03-02 · CVPR 2022 1 · Yao Duan, Chenyang Zhu, Yuqing Lan, Renjiao Yi 외

We introduce Displacement Aware Relation Module (DisARM), a novel neural network module for enhancing the performance of 3D object detection in point cloud scenes. The core idea of our method is that contextual informati…

3D Object Detectionobject-detectionObject DetectionRelation

GA-DAN: Geometry-Aware Domain Adaptation Network for Scene Text Detection and Recognition

2019-07-23 · ICCV 2019 10 · Fangneng Zhan, Chuhui Xue, Shijian Lu

Recent adversarial learning research has achieved very impressive progress for modelling cross-domain data shifts in appearance space but its counterpart in modelling cross-domain shifts in geometry space lags far behind…

Domain AdaptationScene Text DetectionText Detection

Context-Nav: Context-Driven Exploration and Viewpoint-Aware 3D Spatial Reasoning for Instance Navigation

2026-03-10 · Won Shik Jang, Ue-Hwan Kim arxiv

Text-goal instance navigation (TGIN) asks an agent to resolve a single, free-form description into actions that reach the correct object instance among same-category distractors. We present \textit{Context-Nav}, which el…

Spatial Reasoning