Using Artificial Neural Networks to Identify Headings in Newspaper Documents
Document Type
Conference Proceeding
Publication Date
7-20-2003
DOI
https://doi.org/10.1109/IJCNN.2003.1223767
Abstract
Several features for Neural Network based document region identification are tested. Specifically, this paper examines features for headline and subheadline region identification. The Neural Network based region identification algorithm is a key component of a document recognition system that segments a document into regions, classifies them into text, graphic, photo, and other region types, and then uses this classification to guide the processing and analysis of the image. The input data are unusually challenging: low quality images of newspaper documents obtained from microfilmed archives. Experiments on several newspaper documents show that the features used are capable of robust and accurate headline identification.
Publication Information
Zhang, Wei and Andersen, Tim. (2003). "Using Artificial Neural Networks to Identify Headings in Newspaper Documents". Proceedings of the International Joint Conference on Neural Networks, 32283-2287.