Skip to content
Advertisement

Text-to-Speech

312 results

AudioMultimodalText

Irodori-Ja-Spk2-10k

SynDataLab/Irodori-Ja-Spk2-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…

10K–100K·CC-BY-NC-SA·Parquet
AudioMultimodalText

French_game_voice

French Game Voice Dataset Dataset of 100k+ cleaned audio samples of French video game voices with transcriptions. Features…

100K–1M·Parquet
AudioMultimodalText

irodori-refs-10k-v2

Irodori TTS Reference Voices v2 (10K) 10,000 reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign (no ref=True) using a richer…

10K–100K·Apache-2.0·Parquet
AudioMultimodalText

IndicTTS_Bengali

Bengali Indic TTS Dataset This dataset is derived from the Indic TTS Database project, specifically using the Bengali…

10K–100K·CC-BY·Parquet
ImageMultimodalText

emova-sft-speech-231k

EMOVA-SFT-Speech-231K 🤗 EMOVA-Models 🤗 EMOVA-Datasets 🤗 EMOVA-Demo 📄 Paper 🌐 Project-Page 💻 Github 💻 EMOVA-Speech-Tokenizer-Github Overview…

100K–1M·Apache-2.0·Parquet
AudioMultimodalText

Habibi

A systematic and standardized benchmark for the multi-dialect Arabic zero-shot TTS task Paper: "Habibi: Laying the Open-Source Foundation…

10K–100K·Apache-2.0·Parquet
Advertisement