Datasets

Dataset

odia-ocr-synth-v2

by Pritosh

Odia OCR Synthetic Training Data v2 200,000 synthetic image-text pairs for training OCR models on Odia (ଓଡ଼ିଆ) script. Major upgrade over v1 with 29 font variants, realistic backgrounds, and diverse text types.

MITimage-to-text 31 0100K < n < 1M rows

published 27 Mar 2026

Listed inAwesome-Odia-AI

Preview

The Hugging Face dataset viewer has no preview for this one.

Cite this

MITCite

Tags

task_categories:image-to-textlanguage:orlicense:mitsize_categories:100K<n<1Mregion:usocrodiaoriyasynthetictext-recognition