Skew detection and block classification of printed documents
Since the number of daily-received paper-based office documents is overwhelming, the development of document image analysis, which converts the paper-based documents into electronic forms becomes increasingly important. This paper describes a skew detection method which first smoothes the black runs...
Gespeichert in:
Veröffentlicht in: | Image and vision computing 2001-05, Vol.19 (8), p.567-579 |
---|---|
1. Verfasser: | |
Format: | Artikel |
Sprache: | eng |
Schlagworte: | |
Online-Zugang: | Volltext |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Zusammenfassung: | Since the number of daily-received paper-based office documents is overwhelming, the development of document image analysis, which converts the paper-based documents into electronic forms becomes increasingly important. This paper describes a skew detection method which first smoothes the black runs and locates the black–white transitions to emphasize the text lines. Then the skew angle is determined by an improved Hough transform. For the block classification step, a rule-based classifier is presented. The classification rules are derived from the gray level entropy, block aspect ratio, and run length analysis. To evaluate the performance of the proposed methods, a test set of 100 different documents is used. The results of the experiments reveal that all of the 100 documents are successfully skew-corrected and the precision rate and the recall rate of the proposed block classifier are satisfactory. |
---|---|
ISSN: | 0262-8856 1872-8138 |
DOI: | 10.1016/S0262-8856(00)00098-6 |