JOURNAL ARTICLE

Gaussian Constrained Attention Network for Scene Text Recognition

Abstract

Scene text recognition has been a hot topic in computer vision. Recent methods adopt the attention mechanism for sequence prediction which achieve convincing results. However, we argue that the existing attention mechanism faces the problem of attention diffusion, in which the model may not focus on a certain character area. In this paper, we propose Gaussian Constrained Attention Network to deal with this problem. It is a 2D attention-based method integrated with a novel Gaussian Constrained Refinement Module, which predicts an additional Gaussian mask to refine the attention weights. Different from adopting an additional supervision on the attention weights simply, our proposed method introduces an explicit refinement. In this way, the attention weights will be more concentrated and the attention-based recognition network achieves better performance. The proposed Gaussian Constrained Refinement Module is flexible and can be applied to existing attention-based methods directly. The experiments on several benchmark datasets demonstrate the effectiveness of our proposed method. Our code has been available at https://github.com/Pay20Y/GCAN.

Keywords:
Computer science Gaussian Benchmark (surveying) Focus (optics) Code (set theory) Attention network Artificial intelligence Source code Sequence (biology) Gaussian process Machine learning Pattern recognition (psychology) Algorithm Theoretical computer science

Metrics

26
Cited By
2.25
FWCI (Field Weighted Citation Impact)
98
Refs
0.89
Citation Normalized Percentile
Is in top 1%
Is in top 10%

Citation History

Topics

Handwritten Text Recognition Techniques
Physical Sciences →  Computer Science →  Computer Vision and Pattern Recognition
Image Retrieval and Classification Techniques
Physical Sciences →  Computer Science →  Computer Vision and Pattern Recognition
Advanced Image and Video Retrieval Techniques
Physical Sciences →  Computer Science →  Computer Vision and Pattern Recognition

Related Documents

JOURNAL ARTICLE

Context Attention Network for Scene Text Recognition

田荣 董

Journal:   Software Engineering and Applications Year: 2023 Vol: 12 (02)Pages: 345-353
JOURNAL ARTICLE

Orthogonality-constrained multihead self-attention for scene text recognition

Shicheng XuZiqi Zhu

Journal:   Journal of Image and Graphics Year: 2023 Vol: 28 (12)Pages: 3855-3869
JOURNAL ARTICLE

Spatial attention contrastive network for scene text recognition

Fan WangDong Yin

Journal:   Journal of Electronic Imaging Year: 2022 Vol: 31 (04)
© 2026 ScienceGate Book Chapters — All rights reserved.