SKILL.md
Handwriting Recognition Guide
A skill for applying handwriting text recognition (HTR) to digitize historical documents, archival manuscripts, and handwritten research notes. Covers HTR platforms, image preprocessing, model training, post-correction, and integration into digital humanities research workflows.
Handwriting Recognition vs. Printed OCR
Key Differences
Printed Text OCR:
- Characters are standardized and uniform
- Well-solved problem (>99% accuracy on clean scans)
- Tools: Tesseract, ABBYY FineReader, Adobe Acrobat
Handwriting Text Recognition (HTR):
- Characters vary by writer, mood, pen, era
- Much harder -- typically 85-95% character accuracy
- Requires training on specific handwriting styles
- Tools: Transkribus, Kraken, HTR-Flor, Google Cloud Vision
Challenges specific to historical documents:
- Faded ink, bleed-through, stains, tears
- Archaic letterforms and abbreviations
- Multiple hands in one document
- Non-standard orthography
- Mixed languages and scripts
HTR Platforms
Transkribus (State of the Art for Historical Documents)
Pricing note: Transkribus uses a credit-based pricing model. A limited free tier is available, but processing large volumes of pages requires purchasing credits.
Transkribus is the leading platform for historical HTR.
Workflow:
1. Upload document images
2. Automatic layout analysis (detect text regions and baselines)
3. Manual correction of layout (if needed)
4. Apply a pre-trained HTR model (or train your own)
5. Review and correct transcription
6. Export as TEXT, PAGE XML, TEI, DOCX, or PDF
Pre-trained models:
- Noscemus GM (general model for Latin scripts)
- English Writing M1 (18th-19th century English)
- German Kurrent models
- Dutch, French, Italian, Spanish models available
Training a custom model:
- Requires ~15,000-25,000 words of ground truth (manually transcribed)
- Can start with a pre-trained base model and fine-tune
- Training takes 1-8 hours depending on dataset size
Other Tools
| Tool |
|---|
