Skip to content
Advertisement

Text

2,175 results

AudioMultimodalText

VOX-DUB

VOX-DUB is a human-based benchmark for evaluating AI dubbing systems.It includes: Audio fragments with original speech from real…

1K–10K·Parquet
AudioMultimodalText

wuwa-voice-EN

wuwa-voice-EN wuwa voice EN is a dataset of voice line from Wuthering Waves Attribute Value Language English Total…

10K–100K·Parquet
AudioMultimodalText

InstructTTSEval

InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex…

1K–10K·MIT·Parquet
AudioMultimodalText

live-atc-europe

Live ATC Europe — audio + transcriptions Enregistrements live d'air traffic control (ATC) européen, capturés depuis LiveATC.net, segmentés…

100K–1M·Custom / Research-only·Parquet
AudioMultimodalText

Irodori-Ja-Spk1-10k

SynDataLab/Irodori-Ja-Spk1-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…

10K–100K·CC-BY-NC-SA·Parquet
AudioMultimodalText

VieNeu-TTS-140h

pnnbao-ump/VieNeu-TTS-140h Mô tả Dataset A high-quality Vietnamese Text-to-Speech (TTS) dataset containing 74,858 audio samples with phonemized…

10K–100K·Apache-2.0·Arrow
AudioMultimodalText

libritts_r

Dataset Card for LibriTTS-R LibriTTS-R 1 is a sound quality improved version of the LibriTTS corpus ( which…

100K–1M·CC-BY·Parquet
AudioMultimodalText

Quran-Recitations

Quran-Recitations Dataset Overview The Quran-Recitations dataset is a rich and reverent collection of Quranic verses, meticulously paired with…

100K–1M·Parquet
Advertisement