ocular

Historical OCR software

An OCR system designed to transcribe historical documents with high accuracy, handling various challenges such as font variation and code-switching.

Ocular is a state-of-the-art historical OCR system.

GitHub

256 stars
32 watching
48 forks
Language: Java
last commit: over 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
chreul/ocr_testdata_earlyprintedbooksProvides test data and models for training Optical Character Recognition (OCR) systems on historical printed books.10
ocr4all/ocr4allProvides OCR services for historical documents through an intuitive web interface244
ryanfb/ancientgreekocr-ocr-evaluation-toolsA collection of tools and scripts to evaluate the accuracy of Optical Character Recognition (OCR) systems22
ibm/max-ocrAn optical character recognition system deployed as a web service using a trained Tesseract OCR model47
ncsu-libraries/ocracokeA Rails application that enables the creation of OCR capabilities for indexing text from page images and providing search results in IIIF format.34
r1me/ttesseractocr4An Object Pascal binding for the Tesseract OCR engine to perform optical character recognition145
antoniogarrote/clj-tesseractA Clojure wrapper for the Tesseract OCR software, allowing developers to easily integrate optical character recognition capabilities into their applications.54
ponteineptique/toebler-ocrAn OCR project using historical French book data to train models and generate transcriptions.1
hamdikahloun/windows_ocrAn OCR library allowing developers to embed high-quality character recognition functionality in their products.18
mittagessen/krakenAn OCR system optimized for historical and non-Latin scripts, providing layout analysis, character recognition, and support for various formats.757
ocropus/hocr-toolsTools for manipulating and analyzing multi-lingual OCR results by representing them in a standard HTML format373
jzarca01/whofferA React Native app that uses OCR to claim rewards by deciphering images0
tesseract-ocr/docsA collection of documents detailing various aspects and improvements to the Tesseract OCR engine262
ub-mannheim/ocr-gt-toolsA web-based tool for editing and annotating OCR transcriptions of scanned text48
openseg-group/openseg.pytorchProvides a PyTorch implementation of several computer vision tasks including object detection, segmentation and parsing.1,191