AON: Towards Arbitrarily-Oriented Text Recognition

Cheng, Zhanzhan; Xu, Yangliu; Bai, Fan; Niu, Yi; Pu, Shiliang; Zhou, Shuigeng

Computer Science > Computer Vision and Pattern Recognition

arXiv:1711.04226 (cs)

[Submitted on 12 Nov 2017 (v1), last revised 22 Mar 2018 (this version, v2)]

Title:AON: Towards Arbitrarily-Oriented Text Recognition

Authors:Zhanzhan Cheng, Yangliu Xu, Fan Bai, Yi Niu, Shiliang Pu, Shuigeng Zhou

View PDF

Abstract:Recognizing text from natural images is a hot research topic in computer vision due to its various applications. Despite the enduring research of several decades on optical character recognition (OCR), recognizing texts from natural images is still a challenging task. This is because scene texts are often in irregular (e.g. curved, arbitrarily-oriented or seriously distorted) arrangements, which have not yet been well addressed in the literature. Existing methods on text recognition mainly work with regular (horizontal and frontal) texts and cannot be trivially generalized to handle irregular texts. In this paper, we develop the arbitrary orientation network (AON) to directly capture the deep features of irregular texts, which are combined into an attention-based decoder to generate character sequence. The whole network can be trained end-to-end by using only images and word-level annotations. Extensive experiments on various benchmarks, including the CUTE80, SVT-Perspective, IIIT5k, SVT and ICDAR datasets, show that the proposed AON-based method achieves the-state-of-the-art performance in irregular datasets, and is comparable to major existing methods in regular datasets.

Comments:	Accepted by CVPR2018
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1711.04226 [cs.CV]
	(or arXiv:1711.04226v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1711.04226

Submission history

From: Zhanzhan Cheng [view email]
[v1] Sun, 12 Nov 2017 03:11:25 UTC (305 KB)
[v2] Thu, 22 Mar 2018 06:15:45 UTC (335 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2017-11

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Zhanzhan Cheng
Xuyang Liu
Fan Bai
Yi Niu
Shiliang Pu

…

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:AON: Towards Arbitrarily-Oriented Text Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:AON: Towards Arbitrarily-Oriented Text Recognition

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators