paper-with-me

Papers

Sequential Visual and Semantic Consistency for Semi-supervised Text Recognition

2024-02-24 · Mingkun Yang, Biao Yang, Minghui Liao, Yingying Zhu, Xiang Bai

Scene text recognition (STR) is a challenging task that requires large-scale annotated data for training. However, collecting and labeling real text images is expensive and time-consuming, which limits the availability of real data. Therefore, most existing STR methods resort to synthetic data, which may introduce domain discrepancy and degrade the performance of STR models. To alleviate this problem, recent semi-supervised STR methods exploit unlabeled real data by enforcing character-level consistency regularization between weakly and strongly augmented views of the same image. However, these methods neglect word-level consistency, which is crucial for sequence recognition tasks. This paper proposes a novel semi-supervised learning method for STR that incorporates word-level consistency regularization from both visual and semantic aspects. Specifically, we devise a shortest path alignment module to align the sequential visual features of different views and minimize their distance. Moreover, we adopt a reinforcement learning framework to optimize the semantic similarity of the predicted strings in the embedding space. We conduct extensive experiments on several standard and challenging STR benchmarks and demonstrate the superiority of our proposed method over existing semi-supervised STR methods.

📄 PDF Abstract BibTeX arXiv:2402.15806

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Text RecognitionSemantic SimilaritySemantic Textual Similarity

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Semi-Supervised Learning for Visual Bird's Eye View Semantic Segmentation

2023-08-28 · Junyu Zhu, Lina Liu, Yu Tang, Feng Wen 외

Visual bird's eye view (BEV) semantic segmentation helps autonomous vehicles understand the surrounding environment only from images, including static elements (e.g., roads) and dynamic elements (e.g., vehicles, pedestri…

Autonomous VehiclesBird's-Eye View Semantic SegmentationData AugmentationSegmentation+1

Structured Consistency Loss for semi-supervised semantic segmentation

2020-01-14 · Jongmok Kim, Jooyoung Jang, Hyunwoo Park, Seongah Jeong

The consistency loss has played a key role in solving problems in recent studies on semi-supervised learning. Yet extant studies with the consistency loss are limited to its application to classification tasks; extant st…

General ClassificationSegmentationSemantic SegmentationSemi-Supervised Semantic Segmentation

Region-level Contrastive and Consistency Learning for Semi-Supervised Semantic Segmentation

2022-04-28 · Jianrong Zhang, Tianyi Wu, Chuanghao Ding, Hongwei Zhao 외

Current semi-supervised semantic segmentation methods mainly focus on designing pixel-level consistency and contrastive regularization. However, pixel-level regularization is sensitive to noise from pixels with incorrect…

SegmentationSemantic SegmentationSemi-Supervised Semantic Segmentation

Semi-Supervised Semantic Segmentation with High- and Low-level Consistency

2019-08-15 · Sudhanshu Mittal, Maxim Tatarchenko, Thomas Brox

The ability to understand visual information from limited labeled data is an important aspect of machine learning. While image-level classification has been extensively studied in a semi-supervised setting, dense pixel-l…

ClassificationGeneral ClassificationSegmentationSemantic Segmentation+2

CO2: Consistent Contrast for Unsupervised Visual Representation Learning

2020-10-05 · ICLR 2021 1 · Chen Wei, Huiyu Wang, Wei Shen, Alan Yuille

Contrastive learning has been adopted as a core method for unsupervised visual representation learning. Without human annotation, the common practice is to perform an instance discrimination task: Given a query image cro…

Contrastive Learningimage-classificationImage Classificationobject-detection+4