A Text Attention Network for Spatial Deformation Robust Scene Text Image Super-resolution
Jianqi Ma, Zhetong Liang, Lei Zhang
2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Abstract
Scene text image super-resolution aims to increase the resolution and readability of the text in low-resolution images. Though significant improvement has been achieved by deep convolutional neural networks (CNNs), it remains difficult to reconstruct high-resolution images for spatially deformed texts, especially rotated and curve-shaped ones. This is because the current CNN-based methods adopt locality-based operations, which are not effective to deal with the variation caused by deformations. In this paper, we propose a CNN based Text ATTention network (TATT) to address this problem. The semantics of the text are firstly extracted by a text recognition module as text prior information. Then we design a novel transformer-based module, which leverages global attention mechanism, to exert t