paper-with-me

Papers

Reading and Writing: Discriminative and Generative Modeling for Self-Supervised Text Recognition

2022-07-01 · Mingkun Yang, Minghui Liao, Pu Lu, Jing Wang, Shenggao Zhu, Hualin Luo, Qi Tian, Xiang Bai

Existing text recognition methods usually need large-scale training data. Most of them rely on synthetic training data due to the lack of annotated real images. However, there is a domain gap between the synthetic data and real data, which limits the performance of the text recognition models. Recent self-supervised text recognition methods attempted to utilize unlabeled real images by introducing contrastive learning, which mainly learns the discrimination of the text images. Inspired by the observation that humans learn to recognize the texts through both reading and writing, we propose to learn discrimination and generation by integrating contrastive learning and masked image modeling in our self-supervised method. The contrastive learning branch is adopted to learn the discrimination of text images, which imitates the reading behavior of humans. Meanwhile, masked image modeling is firstly introduced for text recognition to learn the context generation of the text images, which is similar to the writing behavior. The experimental results show that our method outperforms previous self-supervised text recognition methods by 10.2%-20.2% on irregular scene text recognition datasets. Moreover, our proposed text recognizer exceeds previous state-of-the-art text recognition methods by averagely 5.3% on 11 benchmarks, with similar model size. We also demonstrate that our pre-trained model can be easily applied to other text-related tasks with obvious performance gain. The code is available at https://github.com/ayumiymk/DiG.

📄 PDF Abstract BibTeX arXiv:2207.00193

Code (1)

ayumiymk/DiG 공식 구현 pytorch

Tasks

Contrastive LearningScene Text Recognition

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

LiteraryTaste: A Preference Dataset for Creative Writing Personalization

2025-11-12 · John Joon Young Chung, Vishakh Padmakumar, Melissa Roemmele, Yi Wang 외 arxiv

People have different creative writing preferences, and large language models (LLMs) for these tasks can benefit from adapting to each user's preferences. However, these models are often trained over a dataset that consi…

Experimental Evidence on Negative Impact of Generative AI on Scientific Learning Outcomes

2023-09-23 · Qirui Ju

In this study, I explored the impact of Generative AI on learning efficacy in academic reading materials using experimental methods. College-educated participants engaged in three cycles of reading and writing tasks. Aft…

Introspective Generative Modeling: Decide Discriminatively

2017-04-25 · Justin Lazarow, Long Jin, Zhuowen Tu

We study unsupervised learning by developing introspective generative modeling (IGM) that attains a generator using progressively learned deep convolutional neural networks. The generator is itself a discriminator, capab…

General Classification

Incentives shape how humans co-create with generative AI

2026-04-04 · Nathanael Jo, Manish Raghavan arxiv

Generative AI is quickly becoming an integral part of people's everyday workflows. Early evidence has shown that while generative AI can increase individual-level productivity, it does so at the cost of collective divers…

Comparing human and LLM proofreading in L2 writing: Impact on lexical and syntactic features

2025-06-10 · Hakyung Sung, Karla Csuros, Min-Chang Sung

This study examines the lexical and syntactic interventions of human and LLM proofreading aimed at improving overall intelligibility in identical second language writings, and evaluates the consistency of outcomes across…

Sentence