Datasets
Dataset
leyu-amharic-speech-corpus-2026
by SoundWaveET
🇪🇹 Leyu Ethiopian Multilingual Speech Corpus 2026 Official research-grade multilingual speech dataset submitted for the Leyu Platform Data Collection Competition 2026, organized by gheero (Leyu Platform Team).
CC-BY-4.0automatic-speech-recognition 179 01K < n < 10K rows
published 6 Aug 2026
Preview
Cannot load the dataset split (in streaming mode) to extract the first rows.
Cite this
CC-BY-4.0audioCite
Tags
task_categories:automatic-speech-recognitionlanguage:amlanguage:orlanguage:tilanguage:sidlanguage:wallicense:cc-by-4.0size_categories:1K<n<10Kmodality:audioregion:usspeechaudioethiopian-languagesamharicafaan-oromotigrinyasidamawolayttaleyu-data-collection-2026