Text-to-Speech
312 results
KYNMAW — Khasi Bhashini Multi-Task Dataset
KYNMAW — Khasi Bhashini Multi-Task Dataset
Irodori-Ja-Spk2-10k
SynDataLab/Irodori-Ja-Spk2-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…
French_game_voice
French Game Voice Dataset Dataset of 100k+ cleaned audio samples of French video game voices with transcriptions. Features…
irodori-refs-10k-v2
Irodori TTS Reference Voices v2 (10K) 10,000 reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign (no ref=True) using a richer…
bengali-tts-missing-v1
Bengali TTS — Missing Rows This dataset contains the rows from rwd51/bengali-tts-combined that are not present in the…
IndicTTS_Bengali
Bengali Indic TTS Dataset This dataset is derived from the Indic TTS Database project, specifically using the Bengali…
ivrit.ai – Knesset Plenums Whisper Training
ivrit.ai - Knesset Plenums Whisper Training
Bambara-ASR-All Audio Dataset
Bambara-ASR-All Audio Dataset
Eka Medical Asr Sample Noise Eval Dataset
Eka Medical Asr Sample Noise Eval Dataset
emova-sft-speech-231k
EMOVA-SFT-Speech-231K 🤗 EMOVA-Models 🤗 EMOVA-Datasets 🤗 EMOVA-Demo 📄 Paper 🌐 Project-Page 💻 Github 💻 EMOVA-Speech-Tokenizer-Github Overview…
Ghana TTS Navigation Corpus (Twi)
Ghana TTS Navigation Corpus (Twi)
Habibi
A systematic and standardized benchmark for the multi-dialect Arabic zero-shot TTS task Paper: "Habibi: Laying the Open-Source Foundation…
InfoRe Technology public dataset №1
InfoRe Technology public dataset №1