Results 261 to 270 of about 44,424,631 (291)
Some of the next articles are maybe not open access.
Fast text line extraction in document images
2012 19th IEEE International Conference on Image Processing, 2012This paper proposes an algorithm for fast text line extraction in document image. Instead of binarization or multi-oriented Gaussian blurring of an image as in the conventional methods, we use integral image and design filters that are proper to detect text regions on the integral image.
Seong Jong Ha, Bora Jin, Nam Ik Cho
openaire +2 more sources
A Text Line Extraction Method for Archival Document Transcription
2020 17th International Multi-Conference on Systems, Signals & Devices (SSD), 2020In order to reinforce the enrichment and exploitation of archival collections, a growing need for computer-aided tools able to assist researchers, historians and archivists in historical document image transcription has been recently highlighted. However, to ensure an efficient text transcription from archival handwritten and printed document images, a
Olfa Mechi +3 more
openaire +1 more source
Text Line Extraction in Handwritten Historical Documents
2017We present a novel approach for the extraction of text lines in handwritten documents using a Convolutional Neural Network to label document image patches as text lines or separators. We first process the document to identify the most suitable patch size on the basis of an overall text line distance estimation.
Capobianco, Samuele, Marinai, Simone
openaire +2 more sources
Ancient document analysis based on text line extraction
2008 19th International Conference on Pattern Recognition, 2008In order to preserve our cultural heritage and for automated document processing libraries and national archives have started digitizing historical documents. In the case of degraded manuscripts (e.g. by mold, humidity, bad storage conditions) the text or parts of it can disappear.
Florian Kleber +3 more
openaire +1 more source
Lines segmentation and word extraction of Arabic handwritten text
Proceedings of the 3rd International Conference on Smart City Applications, 2018Words are often a succession of sub-words (characters, connected components) separated by spaces, in Arabic handwritten its spaces are divided into two types: the first type represents the spaces that separate two connected components of the same word (within-word).
Asmae Lamsaf +3 more
openaire +2 more sources
Extraction of text lines and text blocks on document images based on statistical modeling
International Journal of Imaging Systems and Technology, 1996In this article, we developed a Bayesian model to characterize text line and text block structures on document images using the text word bounding boxes. We posed the extraction problem as finding the text lines and text blocks that maximize the Bayesian probability of the text lines and text blocks given the text word bounding boxes. In particular, we
Su S. Chen +2 more
openaire +2 more sources
Cooperative text and line-art extraction from a topographic map
Proceedings of the Fifth International Conference on Document Analysis and Recognition. ICDAR '99 (Cat. No.PR00318), 1999The black layer is digitized from a USGS topographic map digitized at 1000 dpi. The connected components of this layer are analyzed and separated into line art, text, and icons in two passes. The paired street casings are converted to polylines by vectorization and associated with street labels from the character recognition phase.
Luyang Li +4 more
openaire +1 more source
A Hough based algorithm for extracting text lines in handwritten documents
Proceedings of 3rd International Conference on Document Analysis and Recognition, 2002The method herein proposed detects text lines on handwritten pages which may include either lines oriented in several directions, erasures, or annotations between main lines. The method has a hypothesis-validation strategy which is iteratively activated until the end of the segmentation is reached.
Laurence Likforman-Sulem +2 more
openaire +1 more source
Tree structure for word extraction from handwritten text lines
Eighth International Conference on Document Analysis and Recognition (ICDAR'05), 2005Word extraction from handwritten text lines usually involves the calculation of a line specific threshold which separates the gaps between words from the gaps inside the words in that line. We show that this approach can be improved if the decision about a gap is not only made in terms of a threshold, but also depends on the context of that gap, i.e ...
Tamás Varga, Horst Bunke
openaire +1 more source
Integrated text and line-art extraction from a topographic map
International Journal on Document Analysis and Recognition, 2000Our proposed approach to text and line-art extraction requires accurately locating a text-string box and identifying external line vectors incident on the box. The results of extrapolating these vectors inside the box are passed to an experimental single-font optical char- acter reader (OCR) program, specically trained for the font used for street ...
Luyang Li +4 more
openaire +1 more source

