paper-with-me

홈 › Papers

Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition

2024-07-08 · Bangbang Zhou, Yadong Qu, Zixiao Wang, Zicheng Li, Boqiang Zhang, Hongtao Xie

Recently, scene text recognition (STR) models have shown significant performance improvements. However, existing models still encounter difficulties in recognizing challenging texts that involve factors such as severely distorted and perspective characters. These challenging texts mainly cause two problems: (1) Large Intra-Class Variance. (2) Small Inter-Class Variance. An extremely distorted character may prominently differ visually from other characters within the same category, while the variance between characters from different classes is relatively small. To address the above issues, we propose a novel method that enriches the character features to enhance the discriminability of characters. Firstly, we propose the Character-Aware Constraint Encoder (CACE) with multiple blocks stacked. CACE introduces a decay matrix in each block to explicitly guide the attention region for each token. By continuously employing the decay matrix, CACE enables tokens to perceive morphological information at the character level. Secondly, an Intra-Inter Consistency Loss (I^2CL) is introduced to consider intra-class compactness and inter-class separability at feature space. I^2CL improves the discriminative capability of features by learning a long-term memory unit for each character category. Trained with synthetic data, our model achieves state-of-the-art performance on common benchmarks (94.1% accuracy) and Union14M-Benchmark (61.6% accuracy). Code is available at https://github.com/bang123-box/CFE.

📄 PDF Abstract BibTeX arXiv:2407.05562

Code (1)

bang123-box/cfe 공식 구현 pytorch

Tasks

Scene Text Recognition

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

DenseRAN for Offline Handwritten Chinese Character Recognition

2018-08-13 · Wenchao Wang, Jianshu Zhang, Jun Du, Zi-Rui Wang 외

Recently, great success has been achieved in offline handwritten Chinese character recognition by using deep learning methods. Chinese characters are mainly logographic and consist of basic radicals, however, previous re…

DecoderOffline Handwritten Chinese Character Recognition

Mean-value exergy modeling of internal combustion engines: characterization of feasible operating regions

2021-06-16 · Gabriele Pozzato, Denise Rizzo, Simona Onori

In this paper, a novel mean-value exergy-based modeling framework for internal combustion engines is developed. The characterization of combustion irreversibilities, thermal exchange between the in-cylinder mixture and t…

Drawing and Recognizing Chinese Characters with Recurrent Neural Network

2016-06-21 · Xu-Yao Zhang, Fei Yin, Yan-Ming Zhang, Cheng-Lin Liu 외

Recent deep learning based approaches have achieved great success on handwriting recognition. Chinese characters are among the most widely adopted writing systems in the world. Previous research has mainly focused on rec…

Handwriting Recognition

CASPER in the Machine: Insights into Character Variety in LLM-Generated Stories

2026-06-21 · Anneliese Brei, Abhisheik Sharma, Nicholas Sanaie, Lu Wang 외 arxiv

As LLM-generated text is increasingly used, especially in fictional domains, we explore how much LLM-generated stories differ from human-written stories. In this work, we focus on characters. We borrow definitions from n…

Models In a Spelling Bee: Language Models Implicitly Learn the Character Composition of Tokens

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Standard pretrained language models operate on sequences of subword tokens without direct access to the characters that compose each token’s string representation. We probe the embedding layer of pretrained language mode…

Language ModelingLanguage Modelling