Skip to content
Advertisement

Audio

514 results

AudioMultimodalText

vaja-thai

Vaja-Thai (วาจา) — Combined Thai TTS Dataset A unified, quality-filtered Thai speech dataset combining multiple sources for Text-to-Speech…

100K–1M·Custom / Research-only·Parquet
AudioMultimodalText

CSEMOTIONS

CSEMOTIONS: High-Quality Mandarin Emotional Speech Dataset Paper Code CSEMOTIONS is a high-quality Mandarin emotional speech dataset designed for…

1K–10K·Apache-2.0·Parquet
AudioMultimodalText

sayoko-tts-corpus

サヨ子 音声コーパス ダウンロード方法 データセットを圧縮したzipファイルを、gdriveに置いています。 import gdown url = " RcBAaAyRwEIOWuTQFetVaMUU" gdown.download( url,…

<1K·CC-BY·Text (raw)
AudioMultimodalText

FalAR-TTS

FalAR-TTS FalAR-TTS is a subset of the FalAR dataset ( tailored for speech synthesis in European Portuguese. To…

10K–100K·CC-BY·Parquet
Audio

french-tts-conversational-dataset

French Conversational TTS Dataset Dataset Description This dataset contains high-fidelity French text-to-speech audio clips generated using Mistral's…

<1K·Apache-2.0·Audio (folder)
Audio

LibriQuote

This repository contains the LibriQuote dataset, a speech dataset of fictional character utterances for expressive zero-shot speech synthesis.…

10M–100M·CC-BY-NC
AudioMultimodalTabular

YodasSpeakerPool

Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…

1K–10K·CC-BY
AudioMultimodalText

ne-tts-ccp

NE-TTS Chakma (ccp) Cleaned TTS dataset for Chakma (ccp), a North East Indian language. Derived from the Vaani…

10K–100K·CC-BY·Parquet
AudioMultimodalText

SaSLaW

This repository contains the data of SaSLaW corpus. You can download it via the following command: huggingface-cli download…

1K–10K·CC-BY-NC·Audio (folder)
Advertisement