Datasets
Dataset
ODEN-speech
by BBSRguy
ODEN‑speech 🗣️🇮🇳 Odia Diverse ENsemble Speech Corpus ODEN‑speech merges eight publicly‑available Odia (ଓଡ଼ିଆ) speech corpora into a single 16 kHz, speaker‑aware, text‑cleaned dataset suitable for ASR, TTS, representation learning and multilingual researc…
CC-BY-4.0other 22 1100K < n < 1M rows
published 28 May 2025
Listed inAwesome-Odia-AI
Preview
The dataset does not exist, or is not accessible without authentication (private or gated). Please check the spelling of the dataset name or retry with authentication.
Cite this
CC-BY-4.0audiotextCite
Tags
source_datasets:common_voice_17source_datasets:ljspeechsource_datasets:librittssource_datasets:vctksource_datasets:indicttssource_datasets:mucssource_datasets:sayantan_odialanguage:orlanguage:enlicense:cc-by-4.0size_categories:100K<n<1Mformat:parquetmodality:audiomodality:textlibrary:datasetslibrary:dasklibrary:mlcroissantlibrary:polarsregion:us