VLSP 2020 – VinAI – ASR challenge dataset
VLSP 2020 - VinAI - ASR challenge dataset
3,216 results
VLSP 2020 - VinAI - ASR challenge dataset
Transatlantic Voice Archive
NE-TTS Chakma (ccp) Cleaned TTS dataset for Chakma (ccp), a North East Indian language. Derived from the Vaani…
LinTO DataSet Audio for Arabic Tunisian A collection of Tunisian dialect audio and its annotations for STT task…
This repository contains the data of SaSLaW corpus. You can download it via the following command: huggingface-cli download…
Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from…
Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set…
agent-sft-stitch-zh-tts Voiced version of voidful/agent-sft-stitch-zh: the STITCH-S spoken chunks synthesized with BlueMagpie-TTS (hung yi lee…
Kasem Speech-Text Parallel Dataset
Irodori TTS Reference Voices (12,160) 12,160 Japanese reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign. field type description…
Ganjoor Persian Poetry Recitations (Full)
Vāgdhenu — Sanskrit Chant Corpus
Dataset Card for MusicCaps Dataset Summary The MusicCaps dataset contains 5,521 music examples, each of which is labeled…
Emolia-Thinking (VoiceNet balanced subset)