Skip to content
Advertisement

Text-to-Speech

312 results

Audio

Urdu-Munch-Lina

Urdu-Munch-Lina Processed version of zuhri025/Urdu-Munch with LinaCodec encoding. Dataset Structure This dataset contains 2 batches of audio data…

10K–100K·MIT
Audio

AISHELL-3

AISHELL-3 is a large-scale and high-fidelity multi-speaker Mandarin speech corpus published by Beijing Shell Shell Technology Co.,Ltd. It…

10K–100K·Apache-2.0·Audio (folder)
AudioMultimodalText

bashkort_tts_dataset

Bashkort TTS Dataset The largest open dataset for speech synthesis in the Bashkir language — featuring multi-speaker recordings…

10K–100K·CC-BY·Parquet
AudioMultimodalText

VOX-DUB

VOX-DUB is a human-based benchmark for evaluating AI dubbing systems.It includes: Audio fragments with original speech from real…

1K–10K·Parquet
AudioMultimodalText

wuwa-voice-EN

wuwa-voice-EN wuwa voice EN is a dataset of voice line from Wuthering Waves Attribute Value Language English Total…

10K–100K·Parquet
AudioMultimodalText

InstructTTSEval

InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex…

1K–10K·MIT·Parquet
AudioMultimodalText

live-atc-europe

Live ATC Europe — audio + transcriptions Enregistrements live d'air traffic control (ATC) européen, capturés depuis LiveATC.net, segmentés…

100K–1M·Custom / Research-only·Parquet
Advertisement