Ir directamente a la navegación principal Ir directamente a la búsqueda Ir directamente al contenido principal

Semantics-based content extraction in typewritten historical documents

A. Antonacopoulos*, D. Karatzas

*Autor correspondiente de este trabajo

Producción científica: Capítulo de libroCapítuloInvestigaciónrevisión exhaustiva

Resumen

This paper presents a flexible approach to extracting content from scanned historical documents using semantic information. The final electronic document is the result of a "digital historical document lifecycle" process, where the expert knowledge of the historian/archivist user is incorporated at different stages. Results show that such a conversion strategy aided by (expert) user-specified semantic information and which enables the processing of individual parts of the document in a specialised way, produces superior (in a variety of significant ways) results than document analysis and understanding techniques devised for contemporary documents.

Idioma originalInglés
Título de la publicación alojadaProceedings of the Eighth International Conference on Document Analysis and Recognition
Páginas48-53
Número de páginas6
DOI
EstadoPublicada - 2005

Serie de la publicación

NombreProceedings of the International Conference on Document Analysis and Recognition, ICDAR
Volumen2005
ISSN (versión impresa)1520-5363

Huella

Profundice en los temas de investigación de 'Semantics-based content extraction in typewritten historical documents'. En conjunto forman una huella única.

Citar esto