EAST: An Efficient and Accurate Scene Text Detector

Xinyu Zhou, Cong Yao, He Wen, Yuzhi Wang, Shuchang Zhou, Weiran He, Jiajun Liang

2017-04-11CVPR 2017 7Curved Text Detection Scene Text Detection Text Detection Optical Character Recognition (OCR)

Paper PDF Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code Code

Abstract

Previous approaches for scene text detection have already achieved promising performances across various benchmarks. However, they usually fall short when dealing with challenging scenarios, even when equipped with deep neural network models, because the overall performance is determined by the interplay of multiple stages and components in the pipelines. In this work, we propose a simple yet powerful pipeline that yields fast and accurate text detection in natural scenes. The pipeline directly predicts words or text lines of arbitrary orientations and quadrilateral shapes in full images, eliminating unnecessary intermediate steps (e.g., candidate aggregation and word partitioning), with a single neural network. The simplicity of our pipeline allows concentrating efforts on designing loss functions and neural network architecture. Experiments on standard datasets including ICDAR 2015, COCO-Text and MSRA-TD500 demonstrate that the proposed algorithm significantly outperforms state-of-the-art methods in terms of both accuracy and efficiency. On the ICDAR 2015 dataset, the proposed algorithm achieves an F-score of 0.7820 at 13.2fps at 720p resolution.

Results

Task	Dataset	Metric	Value	Model
Scene Text Detection	Total-Text	Precision	50	EAST
Scene Text Detection	Total-Text	Recall	36.2	EAST
Scene Text Detection	ICDAR 2015	F-Measure	82.9	PAN
Scene Text Detection	ICDAR 2015	Precision	84	PAN
Scene Text Detection	ICDAR 2015	Recall	81.9	PAN
Scene Text Detection	ICDAR 2015	F-Measure	78.2	EAST + PVANET2x RBOX (single-scale)
Scene Text Detection	ICDAR 2015	Precision	83.6	EAST + PVANET2x RBOX (single-scale)
Scene Text Detection	ICDAR 2015	Recall	73.5	EAST + PVANET2x RBOX (single-scale)
Scene Text Detection	MSRA-TD500	F-Measure	76.08	EAST + PVANET2x
Scene Text Detection	MSRA-TD500	Precision	87.28	EAST + PVANET2x
Scene Text Detection	MSRA-TD500	Recall	67.43	EAST + PVANET2x
Scene Text Detection	COCO-Text	F-Measure	39.45	EAST + VGG16
Scene Text Detection	COCO-Text	Precision	50.39	EAST + VGG16
Scene Text Detection	COCO-Text	Recall	32.4	EAST + VGG16

EAST: An Efficient and Accurate Scene Text Detector

Abstract

Results

Related Papers

EAST: An Efficient and Accurate Scene Text Detector

Abstract

Results

Related Papers