Skip to content
Advertisement

Text-to-Speech

312 results

Audio

emolia-thinking-balanced-buckets

Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that…

100K–1M·CC-BY
AudioMultimodalText

BibleMMS

The Dataset associated with the Paper "Meta Learning Text-to-Speech Synthesis in over 7000 Languages" by Florian Lux, Sarina…

100K–1M·MIT·Parquet
Audio

northern-kurdish-raw-audio

Northern Kurdish Raw Audio Collection Overview This repository contains a large collection of raw Northern Kurdish (Kurmanji Kurdish)…

1K–10K·Audio (folder)
Text

en

WTForbes/en Corpus JSONL files are in corpus dir jsonl/. Audio payloads are stored in audio parquet/ shards. WAV…

1M–10M·Custom / Research-only·JSON
Audio

MMEdit-TestSet

MMEdit Test Set A paired audio editing test set for text-guided audio manipulation evaluation, released with MMEdit. Overview…

1K–10K·Apache-2.0·Audio (folder)
AudioMultimodalText

parczech4speech-segmented

ParCzech4Speech (Sentence-Segmented Variant) Dataset Summary ParCzech4Speech (Sentence-Segmented Variant) is a large-scale Czech speech dataset based…

100K–1M·CC-BY·WebDataset
AudioMultimodalText

INTP

INTP: Intelligibility Preference Speech Dataset We establish a synthetic Intelligibility Preference Speech Dataset (INTP), including about 250K…

100K–1M·CC-BY-NC·Parquet
AudioMultimodalText

Pretraining-V1

Indic TTS Unified v1 A large-scale, unified collection of speech data for text-to-speech (TTS) and speech research. This…

10M–100M·CC-BY·Parquet
Text

LEMAS-Dataset-train

Overview This dataset is part of LEMAS-Project (lemas-project.github.io/LEMAS-Project). It contains a large-scale training set (150k+ hours) and a…

100M–1B·CC-BY·JSON
Advertisement