Skip to content
Advertisement

Datasets

3,216 results

AudioMultimodalText

vieneu-tts-140h-dataset

pnnbao-ump/VieNeu-TTS-140h Mô tả Dataset Dataset tiếng Việt chất lượng cao cho Text-to-Speech (TTS) với 74,858 mẫu audio và…

10K–100K·Apache-2.0·Arrow
AudioMultimodalText

vaja-thai

Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech…

100K–1M·Custom / Research-only·Parquet
AudioMultimodalText

CSEMOTIONS

CSEMOTIONS: High-Quality Mandarin Emotional Speech Dataset Paper Code CSEMOTIONS is a high-quality Mandarin emotional speech dataset designed for…

1K–10K·Apache-2.0·Parquet
AudioMultimodalText

sayoko-tts-corpus

サヨ子 音声コーパス ダウンロード方法 データセットを圧縮したzipファイルを、gdriveに置いています。 import gdown url = " RcBAaAyRwEIOWuTQFetVaMUU" gdown.download( url,…

<1K·CC-BY·Text (raw)
AudioMultimodalText

FalAR-TTS

FalAR-TTS FalAR-TTS is a subset of the FalAR dataset ( tailored for speech synthesis in European Portuguese. To…

10K–100K·CC-BY·Parquet
Audio

french-tts-conversational-dataset

French Conversational TTS Dataset Dataset Description This dataset contains high-fidelity French text-to-speech audio clips generated using Mistral's…

<1K·Apache-2.0·Audio (folder)
Audio

LibriQuote

This repository contains the LibriQuote dataset, a speech dataset of fictional character utterances for expressive zero-shot speech synthesis.…

10M–100M·CC-BY-NC
AudioMultimodalTabular

YodasSpeakerPool

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…

1K–10K·CC-BY
Text

CapSpeech

CapSpeech DataSet used for the paper: CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech Please refer to CapSpeech repo…

10M–100M·CC-BY-NC·Parquet
Advertisement