DE eng

Search in the Catalogues and Directories

Hits 1 – 15 of 15

1
CNN-Based Page Segmentation and Object Classification for Counting Population in Ottoman Archival Documentation
In: Journal of Imaging ; Volume 6 ; Issue 5 (2020)
Abstract: Historical document analysis systems gain importance with the increasing efforts in the digitalization of archives. Page segmentation and layout analysis are crucial steps for such systems. Errors in these steps will affect the outcome of handwritten text recognition and Optical Character Recognition (OCR) methods, which increase the importance of the page segmentation and layout analysis. Degradation of documents, digitization errors, and varying layout styles are the issues that complicate the segmentation of historical documents. The properties of Arabic scripts such as connected letters, ligatures, diacritics, and different writing styles make it even more challenging to process Arabic script historical documents. In this study, we developed an automatic system for counting registered individuals and assigning them to populated places by using a CNN-based architecture. To evaluate the performance of our system, we created a labeled dataset of registers obtained from the first wave of population registers of the Ottoman Empire held between the 1840s and 1860s. We achieved promising results for classifying different types of objects and counting the individuals and assigning them to populated places.
Keyword: Arabic script layout analysis; convolutional neural networks; historical document analysis; page segmentation
URL: https://doi.org/10.3390/jimaging6050032
BASE
Hide details
2
Adaptive Algorithms for Automated Processing of Document Images
Agrawal, Mudit. - 2011
BASE
Show details
3
Portable Language-Independent Adaptive Translation from OCR. Phase 1
In: DTIC (2009)
BASE
Show details
4
PLATO: Portable Language-Independent Adaptive Translation from OCR
In: DTIC (2008)
BASE
Show details
5
A Methodology for End-to-End Evaluation of Arabic Document Image Processing Software
In: DTIC (2006)
BASE
Show details
6
Parsing And Tagging Of Bilingual Dictionary
In: http://www.umiacs.umd.edu/lamp/pubs/TechReports/LAMP_106/LAMP_106.pdf (2003)
BASE
Show details
7
Parsing And Tagging Of Bilingual Dictionary
In: http://www.cs.umd.edu/Library/TRs/CS-TR-4529/CS-TR-4529.pdf (2003)
BASE
Show details
8
PARSING AND TAGGING OF BILINGUAL DICTIONARY
In: http://www.cs.umd.edu/Library/TRs/CS-TR-4529/CS-TR-4529.pdf (2003)
BASE
Show details
9
Parsing and Tagging of Bilingual Dictionary
In: DTIC (2003)
BASE
Show details
10
Parsing and Tagging of Binlingual Dictionary
In: DTIC (2003)
BASE
Show details
11
Text extraction in complex color documents
In: http://ipml.ee.duth.gr/~papamark/color_documents.pdf (2002)
BASE
Show details
12
Performance evaluation of document layout analysis algorithms on the UW data set
In: http://isl.ee.washington.edu/~jliang/Postscript/spie97-1.ps.gz (1997)
BASE
Show details
13
Extraction of Text-Related Features for Condensing Image Documents
In: http://www.parc.xerox.com/istl/members/fchen/././projects/qca/dimsum/spie96dimsum.ps.Z (1996)
BASE
Show details
14
OPTICAL CHARACTER RECOGNITION OF HISTORICAL TEXTS: END-USER FOCUSED RESEARCH FOR SLOVENIAN BOOKS AND NEWSPAPERS FROM THE 18TH AND 19TH CENTURY
In: http://nl.ijs.si/imp/bib/NCD21117.pdf
BASE
Show details
15
G.3 [Probability and Statistics]: Distribution Functions General Terms
In: http://www2009.org/proceedings/pdf/p1165.pdf
BASE
Show details

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
15
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern