Datasets
2,147 results
uyghur-common-voice-tts
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
TTS-Multilingual-Test-Set
Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set…
agent-sft-stitch-zh-tts
agent-sft-stitch-zh-tts Voiced version of voidful/agent-sft-stitch-zh: the STITCH-S spoken chunks synthesized with BlueMagpie-TTS (hung yi lee…
Kasem Speech-Text Parallel Dataset
Kasem Speech-Text Parallel Dataset
irodori-tts-refs-12k
Irodori TTS Reference Voices (12,160) 12,160 Japanese reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign. field type description…
Ganjoor Persian Poetry Recitations (Full)
Ganjoor Persian Poetry Recitations (Full)
Vāgdhenu — Sanskrit Chant Corpus
Vāgdhenu — Sanskrit Chant Corpus
MusicCaps
Dataset Card for MusicCaps Dataset Summary The MusicCaps dataset contains 5,521 music examples, each of which is labeled…
linto-dataset-audio-ar-tn-augmented
LinTO DataSet Audio for Arabic Tunisian Augmented A collection of Tunisian dialect audio and its annotations for STT…
nonverbalspeech38k
🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…
audio_data_russian
Dataset Audio Russian This is a dataset with Russian audio data, split into train for tasks like text-to-speech,…
a novel large-scale Vietnamese speech corpus (LSVSC)
a novel large-scale Vietnamese speech corpus (LSVSC)
SimbaBench_dataset
SibmaBench Data Release & Benchmarking To evaluate your model on SimbaBench across all supported tasks (ASR, TTS, and…
enhanced-audiosnippets-long-2-8M
Enhanced Audiosnippets Long 2.8M Enhanced version of mitermix/audiosnippets long 2 8M with speech enhancement, emotion annotations, speaker…