Detecting text in natural scenes with stroke width transform

AbstractBinarization plays an important role in document analysis and recognition (DAR) systems. In this paper, we present our winning algorithm in ICFHR 2018 competition on handwritten document image binarization (H-DIBCO 2018), which is based on background estimation and energy minimization. First, we adopt mathematical morphological operations to estimate and compensate the document background. It uses a disk-shaped structuring element, whose radius is computed by the minimum entropy-based stroke width transform (SWT). Second, we perform Laplacian energy-based segmentation on the compensated document images. Finally, we implement post-processing to preserve text stroke connectivity and eliminate isolated noise. Experimental results indicate that the proposed method outperforms other state-of-the-art techniques on several public available benchmark datasets.

Download Full-text

Stroke Width Transform for Linear Structure Detection: Application to River and Road Extraction from High-Resolution Satellite Images

Lecture Notes in Computer Science - Image Analysis and Recognition ◽

10.1007/978-3-319-59876-5_67 ◽

2017 ◽

pp. 605-613

Author(s):

Moslem Ouled Sghaier ◽

Imen Hammami ◽

Samuel Foucher ◽

Richard Lepage

Keyword(s):

High Resolution ◽

Satellite Images ◽

Linear Structure ◽

Road Extraction ◽

Stroke Width ◽

Structure Detection ◽

High Resolution Satellite Images ◽

Stroke Width Transform

Download Full-text

Modified Stroke Width Transform for Thai Text Detection

2018 International Conference on Information Technology (InCIT) ◽

10.23919/incit.2018.8584869 ◽

2018 ◽

Author(s):

Taravichet Titijaroonroj

Keyword(s):

Text Detection ◽

Stroke Width ◽

Stroke Width Transform

Download Full-text

Text Detection in Natural Images Using Localized Stroke Width Transform

MultiMedia Modeling - Lecture Notes in Computer Science ◽

10.1007/978-3-319-14445-0_5 ◽

2015 ◽

pp. 49-58 ◽

Cited By ~ 2

Author(s):

Wenyan Dong ◽

Zhouhui Lian ◽

Yingmin Tang ◽

Jianguo Xiao

Keyword(s):

Text Detection ◽

Natural Images ◽

Stroke Width ◽

Stroke Width Transform

Download Full-text

An Improved Stroke Width Transform to Detect Race Bib Numbers

Lecture Notes in Computer Science - Pattern Recognition ◽

10.1007/978-3-319-92198-3_27 ◽

2018 ◽

pp. 267-276

Author(s):

Wellington Moreira de Jesus ◽

Díbio Leandro Borges

Keyword(s):

Stroke Width ◽

Stroke Width Transform

Download Full-text

Text detection via edgeless Stroke Width Transform

2014 International Symposium on Intelligent Signal Processing and Communication Systems (ISPACS) ◽

10.1109/ispacs.2014.7024479 ◽

2014 ◽

Cited By ~ 4

Author(s):

Anhar Risnumawan ◽

Chee Seng Chan

Keyword(s):

Text Detection ◽

Stroke Width ◽

Stroke Width Transform

Download Full-text

Smooth Stroke Width Transform for Text Detection

Artificial Intelligence: Methodology, Systems, and Applications - Lecture Notes in Computer Science ◽

10.1007/978-3-319-44748-3_18 ◽

2016 ◽

pp. 183-191 ◽

Cited By ~ 1

Author(s):

Il-Seok Oh ◽

Jin-Seon Lee

Keyword(s):

Text Detection ◽

Stroke Width ◽

Stroke Width Transform

Download Full-text

Extraction of arbitrary text in natural scene image based on stroke width transform

2014 14th International Conference on Intelligent Systems Design and Applications ◽

10.1109/isda.2014.7066257 ◽

2014 ◽

Author(s):

Jinjuli Jameson ◽

Siti Norul Huda Sheikh Abdullah

Keyword(s):

Natural Scene ◽

Stroke Width ◽

Scene Image ◽

Stroke Width Transform

Download Full-text

An Enhanced MSER Pruning Algorithm for Detection and Localization of Bangla Texts from Scene Images

The International Arab Journal of Information Technology ◽

10.34028/iajit/17/3/11 ◽

2019 ◽

Vol 17 (3) ◽

pp. 375-385

Author(s):

Rashedul Islam ◽

Rafiqul Islam ◽

Kamrul Talukder

Keyword(s):

Color Image ◽

False Positives ◽

Text Recognition ◽

Filtering Method ◽

Stroke Width ◽

Benchmark Database ◽

Maximally Stable Extremal Region ◽

Stroke Width Transform ◽

Detection And Localization ◽

F Measure

Text detection and localization have great importance for content based image analysis and text based image indexing. The efficiency of text recognition depends on the efficiency of text localization. So, the main goal of the proposed method is to detect and localize text regions with high accuracy. To achieve this goal, a new and efficient method has been introduced for localization of Bangla text from scene images. In order to improve precision and recall as well as f-measure, Maximally Stable Extremal Region (MSER) based method along with double filtering techniques have been used. As MSER algorithm generates many false positives, we have introduced double filtering method for removing these false positives to increase the f-measure to a great extent. Our proposed method works at three basic levels. Firstly, MSER regions are generated from the input color image by converting it into gray scale image. Secondly, some heuristic features are used to filter out most of the false positives or non-text regions. Lastly, Stroke Width Transform (SWT) based filtering method is used to filter out remaining non-text regions. Remaining components are then grouped into candidate text regions marked by bounding box over each region. As there is no benchmark database for Bangla text, the proposed method is implemented on our own prepared database consisting of 200 scene images of Bangla texts and has got prominent performance. To evaluate the performance of our proposed approach, we have also tested the proposed method on International Conference on Document Analysis and Recognition( ICDAR) 2013 benchmark database and have got a better result than the related existing methods.

Download Full-text

Scene Text Extraction using Stroke Width Transform

International Journal of Computer Sciences and Engineering ◽

10.26438/ijcse/v6i6.375379 ◽

2018 ◽

Vol 6 (6) ◽

pp. 375-379

Author(s):

K. Esther Amulya ◽

P. Sanoop Kumar

Keyword(s):

Text Extraction ◽

Stroke Width ◽

Scene Text ◽

Stroke Width Transform

Download Full-text