Detecting text in natural scene images with conditional clustering and convolution neural network

Anna Zhu, Guoyou Wang, Yangbo Dong, Brian Kenji Iwana

研究成果: Contribution to journalArticle査読

4 被引用数 (Scopus)

抄録

We present a robust method of detecting text in natural scenes. The work consists of four parts. First, automatically partition the images into different layers based on conditional clustering. The clustering operates in two sequential ways. One has a constrained clustering center and conditional determined cluster numbers, which generate small-size subregions. The other has fixed cluster numbers, which generate full-size subregions. After the clustering, we obtain a bunch of connected components (CCs) in each subregion. In the second step, the convolutional neural network (CNN) is used to classify those CCs to character components or noncharacter ones. The output score of the CNN can be transferred to the postprobability of characters. Then we group the candidate characters into text strings based on the probability and location. Finally, we use a verification step. We choose a multichannel strategy to evaluate the performance on the public datasets: ICDAR2011 and ICDAR2013. The experimental results demonstrate that our algorithm achieves a superior performance compared with the state-of-the-art text detection algorithms.

本文言語英語
論文番号053019
ジャーナルJournal of Electronic Imaging
24
5
DOI
出版ステータス出版済み - 9 1 2015

All Science Journal Classification (ASJC) codes

  • Atomic and Molecular Physics, and Optics
  • Computer Science Applications
  • Electrical and Electronic Engineering

フィンガープリント 「Detecting text in natural scene images with conditional clustering and convolution neural network」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル