Text-to-Speech
312 results
Emolia · Filtered · NanoCodec (FSQ) Tokens
Emolia · Filtered · NanoCodec (FSQ) Tokens
ESpeech datasets annotate by Balalaika
ESpeech datasets annotate by Balalaika
Lahgtna Levantine TTS — Synthetic Levantine Arabic & Code-Switching
Lahgtna Levantine TTS — Synthetic Levantine Arabic & Code-Switching
Barranquenho IPA Pronunciation Dictionary
Barranquenho IPA Pronunciation Dictionary
Urdu-Munch-Lina
Urdu-Munch-Lina Processed version of zuhri025/Urdu-Munch with LinaCodec encoding. Dataset Structure This dataset contains 2 batches of audio data…
Common Voice 13 French (Phonemized & Curated)
Common Voice 13 French (Phonemized & Curated)
bashkort_tts_dataset
Bashkort TTS Dataset The largest open dataset for speech synthesis in the Bashkir language — featuring multi-speaker recordings…
Kahwa Postcast Arabic Speech Dataset
Kahwa Postcast Arabic Speech Dataset
VOX-DUB
VOX-DUB is a human-based benchmark for evaluating AI dubbing systems.It includes: Audio fragments with original speech from real…
wuwa-voice-EN
wuwa-voice-EN wuwa voice EN is a dataset of voice line from Wuthering Waves Attribute Value Language English Total…
Syrian Postcast Arabic Speech Dataset
Syrian Postcast Arabic Speech Dataset
InstructTTSEval
InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex…
live-atc-europe
Live ATC Europe — audio + transcriptions Enregistrements live d'air traffic control (ATC) européen, capturés depuis LiveATC.net, segmentés…
Nagatoro Hayase Voice Dataset (140 Clips)
Nagatoro Hayase Voice Dataset (140 Clips)