Emolia · Filtered · NanoCodec (FSQ) Tokens
Emolia · Filtered · NanoCodec (FSQ) Tokens
3,216 results
Emolia · Filtered · NanoCodec (FSQ) Tokens
VCTK This is a processed clone of the VCTK dataset with leading and trailing silence removed using Silero…
Speech Brain Noise Evaluation Dataset
Eka Medical ASR Evaluation Dataset
OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation…
NVSpeech Dataset Overview The NVSpeech dataset provides extensive annotations of paralinguistic vocalizations for Mandarin Chinese speech, aimed at…
Hawrami Raw Audio Collection Overview This repository contains approximately 500 hours of Hawrami Kurdish raw speech collected from…
Spoken-magpie LLMの日本語Instruction Tuning用データllm-jp/magpie-sft-v1.0をCosyVoice2 TTSを使用して音声化した商用利用可能な日本語の音声言語モデルのSFT用データセットです。 ある程度の話者多様性を持つように生成されています。…
MintTTS Pre-tokenized Audio (somu9/hindi-hq)
Yoruba Speech-Text Parallel Dataset
UniST This dataset contains UniST codec-token training data exported from local metadata and codec results. We train UniSS…
Magpie-Speech-Orpheus-125k A ~125k-sample synthetic speech dataset generated by applying the Magpie instruction-synthesis approach to the Orpheus-TTS…
Southern Kurdish Raw Audio Collection Overview This repository contains approximately 170 hours of Southern Kurdish (SDH) raw speech…