Datasets
2,147 results
Irodori-Ja-Spk1-10k
SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…
libritts_p_dataset_20250821_095157
Contribution This dataset is a processed version of the original LibriTTS-P dataset, optimized for use on the Hugging…
Sunbird African Language Technology (SALT) dataset
Sunbird African Language Technology (SALT) dataset
VieNeu-TTS-140h
pnnbao-ump/VieNeu-TTS-140h Mô tả Dataset A high-quality Vietnamese Text-to-Speech (TTS) dataset containing 74,858 audio samples with phonemized…
WorldAudioNaturalConversations Sample Dataset
WorldAudioNaturalConversations Sample Dataset
Speech DAC Tokens (3 Codebooks)
Speech DAC Tokens (3 Codebooks)
libritts_r
Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…
Quran-Recitations
Quran-Recitations Dataset Overview The Quran-Recitations dataset is a rich and reverent collection of Quranic verses, meticulously paired with…
Multilingual Synthetic TTS (Qwen3)
Multilingual Synthetic TTS (Qwen3)
Ghana TTS Navigation Corpus (Ewe)
Ghana TTS Navigation Corpus (Ewe)
SynParaSpeech
Description Here is the SynParaSpeech dataset. SynParaSpeech is the first automated synthesis framework for constructing large-scale paralinguistic…
Turkish Neural Voice Dataset
Turkish Neural Voice Dataset
Nav-train-gnjr-s-t-t
Nav-train-gnjr-s-t-t Persian speech dataset. Audio resampled to 16000 Hz. Splits Split Samples Total Duration (H:MM:SS) Volume train 160931…
dutch-tts-labeled-complete
Dutch TTS Dataset - Complete Labeled A comprehensive Dutch text-to-speech dataset with 596,508 audio samples totaling 234GB of…
cml-tts-filtered-annotated
Dataset Card for Filtred and annotated CML TTS This dataset is an annotated and filtred version of a…
Korean Single Speaker Speech Dataset
Korean Single Speaker Speech Dataset
Multilingual Speech (Adaption)
Multilingual Speech (Adaption)