Results 211 to 220 of about 142,857 (263)
Some of the next articles are maybe not open access.
Analysis of synthetic document images
Proceedings of the Fifth International Conference on Document Analysis and Recognition. ICDAR '99 (Cat. No.PR00318), 1999Interoperability and reusability are becoming challenging problems in document engineering that cannot be solved satisfactorily with converters and filters. This paper advocates a new approach for logical structure recovery by means of document recognition techniques that operate on synthetic images, such as those rendered by a PostScript interpreter ...
Oliver Hitz, Lyse Robadey, Rolf Ingold
openaire +1 more source
1994
A perfectly parallel thinning algorithm, Y.Y. Zhang and P.S.P. Wang background structure in document images, H.S. Baird analysis of form images, D.-C Wang and S.N. Srihari model-based analysis and understanding of check forms, T.H. Minh and H. Bunke document structures - a survey, Y.Y. Tang and C.Y. Suen automatic input of logic diagrams by recognizing
H Bunke, P S P Wang, H Baird
openaire +2 more sources
A perfectly parallel thinning algorithm, Y.Y. Zhang and P.S.P. Wang background structure in document images, H.S. Baird analysis of form images, D.-C Wang and S.N. Srihari model-based analysis and understanding of check forms, T.H. Minh and H. Bunke document structures - a survey, Y.Y. Tang and C.Y. Suen automatic input of logic diagrams by recognizing
H Bunke, P S P Wang, H Baird
openaire +2 more sources
Script Identification of Document Image Analysis
First International Conference on Innovative Computing, Information and Control - Volume I (ICICIC'06), 2006Script identification prior to OCR is necessary in document image analysis. And each script has unique spatial distribution and visual attribute that make it possible to identify itself from other languages. The key technology of script identification algorithm is to abstract effective measure feature.
Juan Cheng +3 more
openaire +1 more source
Chart analysis and recognition in document images
Proceedings of Sixth International Conference on Document Analysis and Recognition, 2002Hidden Markov models are a probabilistic modeling tool for time series data. It has been successfully applied to many areas, such as speech recognition, hand-written character recognition, etc. In this paper, we present a novel statistical approach using ergodic hidden Markov models to recognize scientific charts.
Yan Ping Zhou, Chew Lim Tan
openaire +1 more source
Image based typographic analysis of documents
Proceedings of 2nd International Conference on Document Analysis and Recognition (ICDAR '93), 2002An approach to image based typographic analysis of documents is provided. The problem requires a spatial understanding of the document layout as well as knowledge of the proper syntax. The system performs a page synthesis from the stream of formatting commands defined in a DVI file.
David S. Doermann, Richard Furuta
openaire +1 more source
Document image analysis for active reading
Proceedings of the 2007 international workshop on Semantically aware document processing and indexing, 2007A huge number of documents that were only available in libraries are now on the web. The web access is a solution to protect the cultural heritage and to facilitate knowledge transmission. Most of these documents are displayed as images of the original paper pages and are indexed by hand.
Claudie Faure, Nicole Vincent
openaire +1 more source
Features for printed document image analysis
Object recognition supported by user interaction for service robots, 2003This paper presents features for text/non-text area separation in printed document images. First, it introduces entropic discrimination, i.e., a simple separation using only one feature. Then, a brief recall on existing texture and geometric discriminant parameters proposed in previous research (2001, 2002) is included.
Jean Duong, Hubert Emptoz, Myriam Côté
openaire +1 more source
Treatment of Diagrams in Document Image Analysis
2000Document image analysis is the study of converting documents from paper form to an electronic form that captures the information content of the document. Necessary processing includes recognition of document layout (to determine reading order, and to distinguish text from diagrams), recognition of text (called Optical Character Recognition, OCR), and ...
Dorothea Blostein +2 more
openaire +1 more source
Handwritten document image segmentation and analysis
Pattern Recognition Letters, 1993Abstract The paper proposes automatic segmentation of handwritten document binary images. The Hough transform (HT) over the source image is applied and then the slope of handwritten rows is estimated by analyzing the parameter plane. The inverse HT is used for ‘cutting’ the image and forming ‘strips’, containing the respective handwritten rows ...
Vladimir A. Shapiro +2 more
openaire +1 more source
XML Data Representation in Document Image Analysis
Ninth International Conference on Document Analysis and Recognition (ICDAR 2007), 2007This paper presents the XML-based formats ALTO, TEI, METS used for Digital Libraries and their interest for data representation in a Document Image Analysis and Recognition (DIAR) process. In the first part we briefly present these formats with focus on their adequacy for structural representation and modeling of DIAR data.
Belaid, Abdel +2 more
openaire +2 more sources

