Redirigiendo al acceso original de articulo en 19 segundos...
Inicio  /  Applied Sciences  /  Vol: 9 Par: 22 (2019)  /  Artículo
ARTÍCULO
TITULO

OCR4all?An Open-Source Tool Providing a (Semi-)Automatic OCR Workflow for Historical Printings

Christian Reul    
Dennis Christ    
Alexander Hartelt    
Nico Balbach    
Maximilian Wehner    
Uwe Springmann    
Christoph Wick    
Christine Grundig    
Andreas Büttner and Frank Puppe    

Resumen

Optical Character Recognition (OCR) on historical printings is a challenging task mainly due to the complexity of the layout and the highly variant typography. Nevertheless, in the last few years, great progress has been made in the area of historical OCR, resulting in several powerful open-source tools for preprocessing, layout analysis and segmentation, character recognition, and post-processing. The drawback of these tools often is their limited applicability by non-technical users like humanist scholars and in particular the combined use of several tools in a workflow. In this paper, we present an open-source OCR software called OCR4all, which combines state-of-the-art OCR components and continuous model training into a comprehensive workflow. While a variety of materials can already be processed fully automatically, books with more complex layouts require manual intervention by the users. This is mostly due to the fact that the required ground truth for training stronger mixed models (for segmentation, as well as text recognition) is not available, yet, neither in the desired quantity nor quality. To deal with this issue in the short run, OCR4all offers a comfortable GUI that allows error corrections not only in the final output, but already in early stages to minimize error propagations. In the long run, this constant manual correction produces large quantities of valuable, high quality training material, which can be used to improve fully automatic approaches. Further on, extensive configuration capabilities are provided to set the degree of automation of the workflow and to make adaptations to the carefully selected default parameters for specific printings, if necessary. During experiments, the fully automated application on 19th Century novels showed that OCR4all can considerably outperform the commercial state-of-the-art tool ABBYY Finereader on moderate layouts if suitably pretrained mixed OCR models are available. Furthermore, on very complex early printed books, even users with minimal or no experience were able to capture the text with manageable effort and great quality, achieving excellent Character Error Rates (CERs) below 0.5%. The architecture of OCR4all allows the easy integration (or substitution) of newly developed tools for its main components by standardized interfaces like PageXML, thus aiming at continual higher automation for historical printings.

 Artículos similares

       
 
Cagri Alperen Inan, Ammar Maoui, Yann Lucas and Joëlle Duplay    
Water resource management scenarios have become more crucial for arid to semi-arid regions. Their application prerequisites rigorous hydrological modelling approaches since data are usually exposed to uncertainties and inaccuracies. In this work, Soil Wa... ver más
Revista: Water

 
António Pedro Branco, Cátia Vaz and Alexandre P. Francisco    
There are several tools available to infer phylogenetic trees, which depict the evolutionary relationships among biological entities such as viral and bacterial strains in infectious outbreaks or cancerous cells in tumor progression trees. These tools re... ver más
Revista: Algorithms

 
David Mattie, Zihang Fang, Emi Takahashi, Lourdes Peña Castillo and Jacob Levman    
Diffusion magnetic resonance imaging (MRI) tractography is a powerful tool for non-invasively studying brain architecture and structural integrity by inferring fiber tracts based on water diffusion profiles. This study provided a thorough set of baseline... ver más
Revista: Information

 
Jonas Blattgerste, Jan Behrends and Thies Pfeiffer    
Mobile Augmented Reality (AR) is a promising technology for educational purposes. It allows for interactive, engaging, and spatially independent learning. While the didactic benefits of AR have been well studied in recent years and commodity smartphones ... ver más
Revista: Information

 
Vijanti Ramautar,Sergio España     Pág. 1 - 29
Assessing business operations? ethical, social, and environmental impacts is a key practice for establishing sustainable development. There is a multitude of methods that describes how to perform such assessments. Often these methods are supported by an ... ver más