Skip to content
Advertisement

Text-to-Speech

312 results

Audio

turkish-tts-combined-raw

Türkçe TTS Birleşik Veri Seti 7 farklı açık kaynak Türkçe TTS veri setinin birleşimi. ~81,500 örnek 24kHz SNAC…

10K–100K·CC-BY-SA
MultimodalTabularText

mls-annotated

Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English…

1M–10M·CC-BY·Parquet
AudioMultimodalText

ne-tts-nnp

NE-TTS Wancho (nnp) Cleaned TTS dataset for Wancho (nnp), a North East Indian language. Derived from the Vaani…

1K–10K·CC-BY·Parquet
AudioMultimodalText

vctk

VCTK This is a processed clone of the VCTK dataset with leading and trailing silence removed using Silero…

10K–100K·CC-BY·Parquet
AudioMultimodalText

OpenDialog

OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation…

100K–1M·CC-BY-NC·WebDataset
AudioMultimodalText

Emilia-NV

NVSpeech Dataset Overview The NVSpeech dataset provides extensive annotations of paralinguistic vocalizations for Mandarin Chinese speech, aimed at…

100K–1M·CC-BY-NC-SA·WebDataset
Audio

hawrami-kurdish-raw-audio

Hawrami Raw Audio Collection Overview This repository contains approximately 500 hours of Hawrami Kurdish raw speech collected from…

<1K·Audio (folder)
AudioMultimodalText

spoken-magpie-ja

Spoken-magpie LLMの日本語Instruction Tuning用データllm-jp/magpie-sft-v1.0をCosyVoice2 TTSを使用して音声化した商用利用可能な日本語の音声言語モデルのSFT用データセットです。 ある程度の話者多様性を持つように生成されています。…

100K–1M·Apache-2.0·Parquet
MultimodalTabularText

UniST

UniST This dataset contains UniST codec-token training data exported from local metadata and codec results. We train UniSS…

10M–100M·CC-BY-NC
Advertisement