Deep neural network with attention model for scene text recognition

Shuohao Li,Jun Lei,Min Tang,Qiang Guo,Jun Zhang

doi:10.1049/iet-cvi.2016.0404

Abstract

The authors present a deep neural network (DNN) with attention model for scene text recognition. The proposed model does not require any segmentation of the input text image. The framework is inspired by the attention model presented recently for speech recognition and image captioning. In the proposed framework, feature extraction, feature attention and sequence recognition are integrated in a jointly trainable network. Compared with previous approaches, the following contributions are mainly made. (i) The attention model is applied into DNN to recognise scene text, and it can effectively solve the sequence recognition problem caused by variable length labels. (ii) Rigorous experiments are performed across a number of challenging benchmarks, including IIIT5K, SVT, ICDAR2003 and ICDAR2013 datasets. Results in experiments show that the proposed model is comparable or better than the state‐of‐the‐art methods. (iii) This model only contains 6.5 million parameters. Compared with other DNN models for scene text recognition, this model has the least number of parameters so far.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Deep neural network with attention model for scene text recognition

Abstract

Talk to us

Similar Papers

More From: IET Computer Vision

Lead the way for us

Journal: IET Computer Vision	Publication Date: Aug 22, 2017
Citations: 9

Similar Papers

An end-to-end model for multi-view scene text recognition
Ayan Banerjee ... Cheng-Lin Liu
Pattern Recognition | VOL. 149
Ayan Banerjee, et. al.Ayan Banerjee ... Cheng-Lin Liu
17 Dec 2023
Pattern Recognition | VOL. 149

Verisimilar Image Synthesis for Accurate Detection and Recognition of Texts in Scenes
Fangneng Zhan ... Chuhui Xue
-
Fangneng Zhan, et. al.Fangneng Zhan ... Chuhui Xue
01 Jan 2018
01 Jan 2018

Memory-Augmented Attention Model for Scene Text Recognition
Cong Wang ... Cheng-Lin Liu
-
Cong Wang, et. al.Cong Wang ... Cheng-Lin Liu
01 Aug 2018
01 Aug 2018

A Metaverse text recognition model based on character-level contrastive learning
Le Sun ... Ghulam Muhammad
Applied Soft Computing | VOL. 149
Le Sun, et. al.Le Sun ... Ghulam Muhammad
30 Oct 2023
Applied Soft Computing | VOL. 149

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Deep neural network with attention model for scene text recognition

Abstract

Talk to us

Similar Papers

More From: IET Computer Vision