Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Mathematics

Kennesaw State University

2018

Articles 1 - 1 of 1

Full-Text Articles in Computer Sciences

Automatic Knowledge Extraction From Ocr Documents Using Hierarchical Document Analysis, Mohammad Masum, Sai Kosaraju, Tanju Bayramoglu, Girish Modgil, Mingon Kang Aug 2018

Automatic Knowledge Extraction From Ocr Documents Using Hierarchical Document Analysis, Mohammad Masum, Sai Kosaraju, Tanju Bayramoglu, Girish Modgil, Mingon Kang

Published and Grey Literature from PhD Candidates

Industries can improve their business efficiency by analyzing and extracting relevant knowledge from large numbers of documents. Knowledge extraction manually from large volume of documents is labor intensive, unscalable and challenging. Consequently, there have been a number of attempts to develop intelligent systems to automatically extract relevant knowledge from OCR documents. Moreover, the automatic system can improve the capability of search engine by providing application-specific domain knowledge. However, extracting the efficient information from OCR documents is challenging due to highly unstructured format. In this paper, we propose an efficient framework for a knowledge extraction system that takes keywords based queries …