RUKOPYS: Ukrainian Handwritten Text Recognition Dataset RUKOPYS (Ukrainian: рукопис — manuscript) is the first large scale open dataset for Ukrainian handwritten text recognition (HTR). It spans over a century of Ukrainian handwriting — from 1920s archival documents to present day school homework — and is designed for end to end document understanding: region detection, type classification, and text transcription. Ukrainian is among the largest Slavic languages (45M+ native speakers) yet had no dedicated open HTR dataset prior to RUKOPYS. Competition: RUKOPYS powers the Handwritten to Data challenge on Kaggle (April 16 — June 15, 2026). Submit your HTR model predictions and compete for $7,000 in prizes. What Makes RUKOPYS Different Most HTR datasets are built from a single source — one archive, one corpus, one handwriting style. RUKOPYS is deliberately the opposite. It combines four sources that differ across every dimension that makes handwriting recognition hard: Dimension Range in RUKOPYS Time period 1919–1935 (archival pen & ink) → 2020–2025 (modern ballpoint, pencil) Writers School children (grades 5–11), university students, adult citizens Document type Archival state documen…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy