Datasets

Dataset

leyu-ethiopian-multilingual-speech-corpus-2026

by SoundWaveET

🇪🇹 Leyu Ethiopian Multilingual Speech Corpus 2026 Official research-grade multilingual speech dataset submitted for the Leyu Platform Data Collection Competition 2026, organized by gheero (Leyu Platform Team).

CC-BY-4.0automatic-speech-recognition 115 01K < n < 10K rows

published 20 Aug 2026

Preview

Cannot load the dataset split (in streaming mode) to extract the first rows.

Cite this

CC-BY-4.0audioCite

Tags

task_categories:automatic-speech-recognitionlanguage:amlanguage:orlanguage:tilanguage:sidlanguage:wallicense:cc-by-4.0size_categories:1K<n<10Kmodality:audioregion:usspeechaudioethiopian-languagesamharicafaan-oromotigrinyasidamawolayttaleyu-data-collection-2026