Text-to-Speech
312 results
emolia-thinking-balanced-buckets
Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that…
viVoice: Enabling Vietnamese Multi-Speaker Speech Synthesis
viVoice: Enabling Vietnamese Multi-Speaker Speech Synthesis
BibleMMS
The Dataset associated with the Paper "Meta Learning Text-to-Speech Synthesis in over 7000 Languages" by Florian Lux, Sarina…
northern-kurdish-raw-audio
Northern Kurdish Raw Audio Collection Overview This repository contains a large collection of raw Northern Kurdish (Kurmanji Kurdish)…
NaturalVoices Restored (16 kHz, Sidon + UTMOS-filtered)
NaturalVoices Restored (16 kHz, Sidon + UTMOS-filtered)
Codemixed_New
Codemixed ASR Dataset Unified collection of code-mixed ASR datasets.
MMEdit-TestSet
MMEdit Test Set A paired audio editing test set for text-guided audio manipulation evaluation, released with MMEdit. Overview…
parczech4speech-segmented
ParCzech4Speech (Sentence-Segmented Variant) Dataset Summary ParCzech4Speech (Sentence-Segmented Variant) is a large-scale Czech speech dataset based…
INTP
INTP: Intelligibility Preference Speech Dataset We establish a synthetic Intelligibility Preference Speech Dataset (INTP), including about 250K…
Pretraining-V1
Indic TTS Unified v1 A large-scale, unified collection of speech data for text-to-speech (TTS) and speech research. This…
LEMAS-Dataset-train
Overview This dataset is part of LEMAS-Project (lemas-project.github.io/LEMAS-Project). It contains a large-scale training set (150k+ hours) and a…