ne-tts-ccp
NE-TTS Chakma (ccp) Cleaned TTS dataset for Chakma (ccp), a North East Indian language. Derived from the Vaani…
1,940 results
NE-TTS Chakma (ccp) Cleaned TTS dataset for Chakma (ccp), a North East Indian language. Derived from the Vaani…
LinTO DataSet Audio for Arabic Tunisian A collection of Tunisian dialect audio and its annotations for STT task…
This repository contains the data of SaSLaW corpus. You can download it via the following command: huggingface-cli download…
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set…
agent-sft-stitch-zh-tts Voiced version of voidful/agent-sft-stitch-zh: the STITCH-S spoken chunks synthesized with BlueMagpie-TTS (hung yi lee…
Kasem Speech-Text Parallel Dataset
Irodori TTS Reference Voices (12,160) 12,160 Japanese reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign. field type description…
Ganjoor Persian Poetry Recitations (Full)
Vāgdhenu — Sanskrit Chant Corpus
Dataset Card for MusicCaps Dataset Summary The MusicCaps dataset contains 5,521 music examples, each of which is labeled…
LinTO DataSet Audio for Arabic Tunisian Augmented A collection of Tunisian dialect audio and its annotations for STT…
🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…
Dataset Audio Russian This is a dataset with Russian audio data, split into train for tasks like text-to-speech,…