FalAR-TTS
FalAR-TTS FalAR-TTS is a subset of the FalAR dataset ( tailored for speech synthesis in European Portuguese. To…
2,175 results
FalAR-TTS FalAR-TTS is a subset of the FalAR dataset ( tailored for speech synthesis in European Portuguese. To…
InfoRe Technology public dataset №2
Mohamed Khairy Arabic Speech Dataset
Hinglish Concatenated Audio Dataset
VoxCPM Ghana — Precomputed AudioVAE Latents
Use this dataset in conjuction with: YodasSpeakerPool YodasSpeakerPool is a curated, richly-annotated multi-speaker dataset featuring 7,600 unique…
OpenSTT annotate by Balalaika
VLSP 2020 - VinAI - ASR challenge dataset
Transatlantic Voice Archive
NE-TTS Chakma (ccp) Cleaned TTS dataset for Chakma (ccp), a North East Indian language. Derived from the Vaani…
LinTO DataSet Audio for Arabic Tunisian A collection of Tunisian dialect audio and its annotations for STT task…
This repository contains the data of SaSLaW corpus. You can download it via the following command: huggingface-cli download…
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set…
agent-sft-stitch-zh-tts Voiced version of voidful/agent-sft-stitch-zh: the STITCH-S spoken chunks synthesized with BlueMagpie-TTS (hung yi lee…
Kasem Speech-Text Parallel Dataset