EAST: An Efficient and Accurate Scene Text Detector

被引:983
作者
Zhou, Xinyu [1 ]
Yao, Cong [1 ]
Wen, He [1 ]
Wang, Yuzhi [1 ]
Zhou, Shuchang [1 ]
He, Weiran [1 ]
Liang, Jiajun [1 ]
机构
[1] Megvii Technol Inc, Beijing, Peoples R China
来源
30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017) | 2017年
关键词
D O I
10.1109/CVPR.2017.283
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Previous approaches for scene text detection have already achieved promising performances across various benchmarks. However, they usually fall short when dealing with challenging scenarios, even when equipped with deep neural network models, because the overall performance is determined by the interplay of multiple stages and components in the pipelines. In this work, we propose a simple yet powerful pipeline that yields fast and accurate text detection in natural scenes. The pipeline directly predicts words or text lines of arbitrary orientations and quadrilateral shapes in full images, eliminating unnecessary intermediate steps (e.g., candidate aggregation and word partitioning), with a single neural network. The simplicity of our pipeline allows concentrating efforts on designing loss functions and neural network architecture. Experiments on standard datasets including ICDAR 2015, COCO-Text and MSRA-TD500 demonstrate that the proposed algorithm significantly outperforms state-of-the-art methods in terms of both accuracy and efficiency. On the ICDAR 2015 dataset, the proposed algorithm achieves an F-score of 0.7820 at 13.2fps at 720p resolution.
引用
收藏
页码:2642 / 2651
页数:10
相关论文
共 48 条
  • [1] [Anonymous], P ECCV
  • [2] [Anonymous], 2015, P CVPR
  • [3] [Anonymous], P ICDAR
  • [4] [Anonymous], 2010, P CVPR
  • [5] [Anonymous], P CVPR
  • [6] [Anonymous], 2016, ARXIV160406646
  • [7] [Anonymous], 2010, P ACCV
  • [8] [Anonymous], 2014, P ECCV
  • [9] [Anonymous], P ICCV
  • [10] [Anonymous], 2012, P CVPR