Character string extraction from newspaper headlines with a background design by recognizing a combination of connected components

Hiroaki Takebe, Yutaka Katsuyama, Satoshi Naoi

研究成果: Conference article査読

2 被引用数 (Scopus)

抄録

In this paper we propose a new method of extracting a character string from images with a background design. In Japanese newspaper headlines, it is common for character components to be placed independent of background components. In view of this, we represent a character string candidate as a consistent combination of connected components, and we calculate its character string resemblance value. In this case, a character string resemblance value of a combination of connected components depends upon its character recognition result and the area of the rectangular area occupied by it. We then extract the combination of connected components that has the maximum character string resemblance value. We applied this method to 142 headline images. The results show that the method accurately extracted a character string from various kinds of images with a background design and the method has a favorable processing speed.

本文言語English
ページ(範囲)22-29
ページ数8
ジャーナルProceedings of SPIE - The International Society for Optical Engineering
3651
DOI
出版ステータスPublished - 1999
外部発表はい
イベントProceedings of the 1999 6th Annual Conference on Document Recognition and Retrieval VI - San Jose, CA, USA
継続期間: 1999 1 271999 1 28

ASJC Scopus subject areas

  • 電子材料、光学材料、および磁性材料
  • 凝縮系物理学
  • コンピュータ サイエンスの応用
  • 応用数学
  • 電子工学および電気工学

フィンガープリント

「Character string extraction from newspaper headlines with a background design by recognizing a combination of connected components」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル