VOX-DUB
VOX-DUB is a human-based benchmark for evaluating AI dubbing systems.It includes: Audio fragments with original speech from real…
1,501 results
VOX-DUB is a human-based benchmark for evaluating AI dubbing systems.It includes: Audio fragments with original speech from real…
wuwa-voice-EN wuwa voice EN is a dataset of voice line from Wuthering Waves Attribute Value Language English Total…
Syrian Postcast Arabic Speech Dataset
InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex…
Live ATC Europe — audio + transcriptions Enregistrements live d'air traffic control (ATC) européen, capturés depuis LiveATC.net, segmentés…
Sagalee: An Open Source ASR Dataset for Oromo Language
SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…
Contribution This dataset is a processed version of the original LibriTTS-P dataset, optimized for use on the Hugging…
Sunbird African Language Technology (SALT) dataset
WorldAudioNaturalConversations Sample Dataset
Speech DAC Tokens (3 Codebooks)
Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…
Quran-Recitations Dataset Overview The Quran-Recitations dataset is a rich and reverent collection of Quranic verses, meticulously paired with…
Multilingual Synthetic TTS (Qwen3)
Description Here is the SynParaSpeech dataset. SynParaSpeech is the first automated synthesis framework for constructing large-scale paralinguistic…
Turkish Neural Voice Dataset
Nav-train-gnjr-s-t-t Persian speech dataset. Audio resampled to 16000 Hz. Splits Split Samples Total Duration (H:MM:SS) Volume train 160931…