Skip to content
Advertisement

Text-to-Speech

312 results

AudioMultimodalText

irodori-tts-refs-12k

Irodori TTS Reference Voices (12,160) 12,160 Japanese reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign. field type description…

10K–100K·Apache-2.0·Parquet
MultimodalTabularText

MusicCaps

Dataset Card for MusicCaps Dataset Summary The MusicCaps dataset contains 5,521 music examples, each of which is labeled…

1K–10K·CC-BY-SA·CSV
AudioMultimodalText

nonverbalspeech38k

🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…

10K–100K·CC-BY-NC·Parquet
AudioMultimodalText

SimbaBench_dataset

SibmaBench Data Release & Benchmarking To evaluate your model on SimbaBench across all supported tasks (ASR, TTS, and…

100K–1M·CC-BY·Parquet
AudioMultimodalText

SpeechJudge-Data

SpeechJudge-Data: A Large-Scale Human Feedback Corpus for Speech Generation Introduction SpeechJudge-Data is a large-scale human feedback corpus of…

10K–100K·CC-BY-NC·Parquet
Advertisement