Skip to content
Advertisement

English

1,233 results

AudioMultimodalText

echo-clones-4m-en

echo-clones-4m-en ~4 M English TTS clone utterances generated with EchoTTS (jordand/echo-tts-base). Sample rate: 44 100 Hz, 16-bit PCM…

1M–10M·Apache-2.0·Parquet
Audio

PersonalHub

Personal Hub: Exploring High-Expressiveness Speech Data through Spatio-Temporal Feature Integration and Model Fine-Tuning Introduction In this work,…

MIT
Audio

DailyTalkContiguous-MoodyGirl

Moody Girl: Emotion-Tagged DailyTalk Dataset This dataset is an emotion-tagged version of the DailyTalk dataset, enhanced with emotion…

CC-BY-SA
AudioMultimodalText

openwhisper

Streaming ASR Dataset This dataset is designed for training real-time (streaming) ASR models, with a focus on handling…

100K–1M·MIT·Parquet
Audio

AIR-Bench-Dataset

AIR-Bench Arxiv: is the AIR-Bench dataset download page.AIR-Bench encompasses two dimensions: foundation and chat benchmarks. The former consists…

<1K·CC-BY-NC·Audio (folder)
Advertisement