4.5 Article

Color Reduction for Complex Document Images

出版社

WILEY
DOI: 10.1002/ima.20174

关键词

color reduction; text information extraction; mean-shift; edge preserving smoothing

向作者/读者索取更多资源

A new technique for color reduction of complex document images is presented in this article. It reduces significantly the number of colors of the document image (less than 15 colors in most of the cases) so as to have solid characters and uniform local backgrounds. Therefore, this technique can be used as a preprocessing step by text information extraction applications. Specifically, using the edge map of the document image, a representative set of samples is chosen that constructs a 3D color histogram. Based on these samples in the 3D color space, a relatively large number of colors (usually no more than 100 colors) are obtained by using a simple clustering procedure. The final colors are obtained by applying a mean-shift based procedure. Also, an edge preserving smoothing filter is used as a preprocessing stage that enhances significantly the quality of the initial image. Experimental results prove the method's capability of producing correctly segmented complex color documents where the character elements can be easily extracted as connected components. (C) 2009 Wiley Periodicals, Inc. Int J Imaging Syst Technol, 19, 14-26, 2009; Published online in Wiley InterScience (www.interscience.wiley.com). DOI 10.1002/ima.20174

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.5
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据