Document Analysis And Classification Based On Passing Window [PDF]
In this paper we present Document analysis and classification system to segment and classify contents of Arabic document images. This system includes preprocessing, document segmentation, feature extraction and document classification.
ZAHER BAMASOOD
doaj
Analysis of Hyperspectral Data to Develop an Approach for Document Images
Hyperspectral data analysis is being utilized as an effective and compelling tool for image processing, providing unprecedented levels of information and insights for various applications.
Zainab Zaman +2 more
doaj +1 more source
Office Documents Classification under Limited Sample. A Case of Table Detection Inside Court Files
Deep convolutional neural networks (CNNs) became an industry standard in image processing. However, in order to keep their high efficiency, a large annotated sample is required in the case of supervised learning.
Paweł Baranowski, Adrian Stepniak
doaj +1 more source
Utilizing image and caption information for biomedical document classification [PDF]
Abstract Motivation Biomedical research findings are typically disseminated through publications. To simplify access to domain-specific knowledge while supporting the research community, several biomedical databases devote significant effort to manual curation of the literature—a labor intensive ...
Pengyuan Li 0001 +9 more
openaire +4 more sources
Analyzing the potential of active learning for document image classification
Abstract Deep learning has been extensively researched in the field of document analysis and has shown excellent performance across a wide range of document-related tasks. As a result, a great deal of emphasis is now being placed on its practical deployment and integration into modern industrial document ...
Saifullah Saifullah +3 more
openaire +1 more source
PHTI: Pashto Handwritten Text Imagebase for Deep Learning Applications
Document Image Analysis (DIA) is one of the research areas of Artificial Intelligence (AI) that converts document images into machine-readable codes. In DIA systems, Optical Character Recognition (OCR) plays a key role in digitizing document images.
Ibrar Hussain +5 more
doaj +1 more source
BANGLA HANDWRITTEN CHARACTER RECOGNITION USING CONVOLUTION NEURAL NETWORK
Since, last one-decade, numerous deep learning models have been designed to resolve handwritten character recognition task in languages, namely, English, Chinese, Arabic, Japanese and Russian.
Shankha De, Arpana Rawal
doaj +1 more source
Tool for Parsing Important Data from Web Pages
This paper discusses the tool for the main text and image extraction (extracting and parsing the important data) from a web document. This paper describes our proposed algorithm based on the Document Object Model (DOM) and natural language processing ...
Martina Radilova +4 more
doaj +1 more source
Semantic Document Image Classification Based on Valuable Text Pattern [PDF]
Knowledge extraction from detected document image is a complex problem in the field of information technology. This problem becomes more intricate when we know, a negligible percentage of the detected document images are valuable.
Hossein Pourghassem +2 more
doaj
Page Layout Analysis of the Document Image Based on the Region Classification in a Decision Hierarchical Structure [PDF]
The conversion of document image to its electronic version is a very important problem in the saving, searching and retrieval application in the official automation system. For this purpose, analysis of the document image is necessary.
Hossein Pourghassem
doaj

