toebler-ocr

Book Transcription Project

An OCR project using historical French book data to train models and generate transcriptions.

GitHub

1 stars
4 watching
0 forks
Language: HTML
last commit: over 7 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
chreul/ocr_testdata_earlyprintedbooksProvides test data and models for training Optical Character Recognition (OCR) systems on historical printed books.10
jbaiter/archiscribeA tool for transcribing OCR data from archival documents17
hamdikahloun/windows_ocrAn OCR library allowing developers to embed high-quality character recognition functionality in their products.18
cpitclaudel/alectryonA tool for processing Coq and Lean 4 code embedded in text documents237
dannnylo/rtesseractA Ruby library providing an interface to the Tesseract OCR system.838
jean-baptiste-camps/froc-mssDevelops models to transcribe handwritten text from Old French and Old Occitan medieval manuscripts0
mittagessen/krakenAn OCR system optimized for historical and non-Latin scripts, providing layout analysis, character recognition, and support for various formats.757
tberg12/ocularAn OCR system designed to transcribe historical documents with high accuracy, handling various challenges such as font variation and code-switching.256
rescribe/carolineminuscule-groundtruthAn OCR ground truth repository for Caroline Minuscule manuscripts.11
ub-mannheim/ocr-gt-toolsA web-based tool for editing and annotating OCR transcriptions of scanned text48
benwbrum/fromthepageA wiki-like application for collaborative transcription of handwritten documents from scanned pages.171
bytedance/piano_transcriptionDevelops software for accurately transcribing piano recordings into MIDI files using machine learning models.1,676
cneud/ocr-conversionA collection of scripts and stylesheets for converting data between different OCR formats.72
carreau/jupyter-bookA project that creates a book on Jupyter, focusing on its capabilities and applications19
nytud/hadifogoly-adatbazisAn attempt to transcribe Cyrillic text into Hungarian script for a large dataset of WWII prisoner-of-war records23