Sagalee: An Open Source ASR Dataset for Oromo Language
Sagalee: An Open Source ASR Dataset for Oromo Language
748 results
Sagalee: An Open Source ASR Dataset for Oromo Language
SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…
Sunbird African Language Technology (SALT) dataset
pnnbao-ump/VieNeu-TTS-140h Mô tả Dataset A high-quality Vietnamese Text-to-Speech (TTS) dataset containing 74,858 audio samples with phonemized…
Multilingual Synthetic TTS (Qwen3)
Description Here is the SynParaSpeech dataset. SynParaSpeech is the first automated synthesis framework for constructing large-scale paralinguistic…
Korean Single Speaker Speech Dataset
Multilingual Speech (Adaption)
Dataset Card for Dataset Name This dataset is collected from youtube.
The "Thorsten-Voice" dataset This truly open source (CC0 license) german (🇩🇪) voice dataset contains about 40 hours of…
Mostafa Mahmoud Arabic Speech Dataset
Deep Confessions Podcast Arabic Speech Dataset
MiscSpeech-ja This dataset comprises audio and corresponding transcripts collected from a diverse range of YouTube videos and Podcasts.
Türkçe TTS Birleşik Veri Seti 7 farklı açık kaynak Türkçe TTS veri setinin birleşimi. ~81,500 örnek 24kHz SNAC…