uyghur-common-voice-tts
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
514 results
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set…
agent-sft-stitch-zh-tts Voiced version of voidful/agent-sft-stitch-zh: the STITCH-S spoken chunks synthesized with BlueMagpie-TTS (hung yi lee…
Kasem Speech-Text Parallel Dataset
Irodori TTS Reference Voices (12,160) 12,160 Japanese reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign. field type description…
Ganjoor Persian Poetry Recitations (Full)
Vāgdhenu — Sanskrit Chant Corpus
Emolia-Thinking (VoiceNet balanced subset)
LinTO DataSet Audio for Arabic Tunisian Augmented A collection of Tunisian dialect audio and its annotations for STT…
🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…
Dataset Audio Russian This is a dataset with Russian audio data, split into train for tasks like text-to-speech,…
a novel large-scale Vietnamese speech corpus (LSVSC)
SibmaBench Data Release & Benchmarking To evaluate your model on SimbaBench across all supported tasks (ASR, TTS, and…