Datasets

Dataset

nirantar

by adjaysagar

Nirantar Nirantar speech dataset (22 Indian languages) in Hugging Face format. Source: AI4Bharat/Nirantar.

CC-BY-4.0automatic-speech-recognition 30 2100K < n < 1M rows

published 8 Feb 2026

Preview

The dataset does not exist, or is not accessible without authentication (private or gated). Please check the spelling of the dataset name or retry with authentication.

Cite this

CC-BY-4.0audiotextCite

Tags

task_categories:automatic-speech-recognitiontask_categories:text-to-speechlanguage:bnlanguage:brxlanguage:doilanguage:gulanguage:hilanguage:knlanguage:koklanguage:kslanguage:mailanguage:mllanguage:mnilanguage:mrlanguage:nelanguage:orlanguage:palanguage:salanguage:satlanguage:sdlanguage:talanguage:telanguage:urlicense:cc-by-4.0size_categories:100K<n<1Mformat:parquetmodality:audiomodality:textlibrary:datasetslibrary:dasklibrary:polarslibrary:mlcroissantregion:us