Datasets
Dataset
odia-ocr-synth-v2
by Pritosh
Odia OCR Synthetic Training Data v2 200,000 synthetic image-text pairs for training OCR models on Odia (ଓଡ଼ିଆ) script. Major upgrade over v1 with 29 font variants, realistic backgrounds, and diverse text types.
MITimage-to-text 31 0100K < n < 1M rows
published 27 Mar 2026
Listed inAwesome-Odia-AI
Preview
The Hugging Face dataset viewer has no preview for this one.
Cite this
MITCite
Tags
task_categories:image-to-textlanguage:orlicense:mitsize_categories:100K<n<1Mregion:usocrodiaoriyasynthetictext-recognition