Abstract:Recently, scene text detection has become an active research topic in computer vision and document analysis, because of its great importance and significant challenge. However, vast majority of the existing methods detect text within local regions, typically through extracting character, word or line level candidates followed by candidate aggregation and false positive elimination, which potentially exclude the effect of wide-scope and long-range contextual cues in the scene. To take full advantage of the rich information available in the whole natural image, we propose to localize text in a holistic manner, by casting scene text detection as a semantic segmentation problem. The proposed algorithm directly runs on full images and produces global, pixel-wise prediction maps, in which detections are subsequently formed. To better make use of the properties of text, three types of information regarding text region, individual characters and their relationship are estimated, with a single Fully Convolutional Network (FCN) model. With such predictions of text properties, the proposed algorithm can simultaneously handle horizontal, multi-oriented and curved text in real-world natural images. The experiments on standard benchmarks, including ICDAR 2013, ICDAR 2015 and MSRA-TD500, demonstrate that the proposed algorithm substantially outperforms previous state-of-the-art approaches. Moreover, we report the first baseline result on the recently-released, large-scale dataset COCO-Text.

Scene Text Detection by Leveraging Multi-Channel Information and Local Context

Text Detection in Natural Scene Images Leveraging Context Information.

Text Detection in Scene Images Based on Exhaustive Segmentation

Robust Text Detection in Natural Scene Images

Text detection approach based on confidence map and context information.

Scene Text Detection Via Extremal Region Based Double Threshold Convolutional Network Classification.

Scene Text Detection via Holistic, Multi-Channel Prediction

A Cascaded Method for Text Detection in Natural Scene Images

Robust Scene Text Detection Based on Color Consistency

Text Detection in Natural Scene Images Based on Color Prior Guided MSER.

Scene Text Detection Using Adaptive Color Reduction, Adjacent Character Model and Hybrid Verification Strategy

Multi-orientation Scene Text Detection Leveraging Background Suppression

Scene Text Recognition from Two-Dimensional Perspective

Leveraging Surrounding Context for Scene Text Detection

Natural Scene Text Detection with Multi-Channel Connected Component Segmentation

Multi-oriented Scene Text Detection via Corner Localization and Region Segmentation

Natural scene text detection by multi-scale adaptive color clustering and non-text filtering.

Natural Scene Text Detection with Multi-Layer Segmentation and Higher Order Conditional Random Field Based Analysis.

Integrated Method for Text Detection in Natural Scene Images

Scene Text Detection Method Based on the Hierarchical Model

Text Detection and Recognition in Natural Scene with Edge Analysis