Datasets

Dataset

nirantar

by shiprocket-ai

Nirantar Nirantar speech dataset (22 Indian languages) in Hugging Face format. Source: AI4Bharat/Nirantar.

CC-BY-4.0automatic-speech-recognition 26 2100K < n < 1M rows

published 8 Feb 2026

Preview

The dataset does not exist, or is not accessible without authentication (private or gated). Please check the spelling of the dataset name or retry with authentication.

Cite this

CC-BY-4.0audiotextCite

Tags

task_categories:automatic-speech-recognitiontask_categories:text-to-speechlanguage:bnlanguage:brxlanguage:doilanguage:gulanguage:hilanguage:knlanguage:koklanguage:kslanguage:mailanguage:mllanguage:mnilanguage:mrlanguage:nelanguage:orlanguage:palanguage:salanguage:satlanguage:sdlanguage:talanguage:telanguage:urlicense:cc-by-4.0size_categories:100K<n<1Mformat:parquetmodality:audiomodality:textlibrary:datasetslibrary:dasklibrary:polarslibrary:mlcroissantregion:us