2 research outputs found

    Features for Neural Net Based Region Identification of Newspaper Documents

    No full text
    Several features for neural network based document region identification are tested. Specifically, this paper examines features for non-text region identification. The neural network based region identification algorithm is a key component of a document recognition system that segments a document into regions, classifies them into text, graphic, photo, and other region types, and then uses this classification to guide the processing and analysis of the image. The input data are unusually challenging: low quality images of newspaper documents obtained from microfilmed archives. The results compare favorably with other results reported in the literature
    corecore