dowis
Do What I Say (DOWIS): A Spoken Prompt Dataset for Instruction-Following NEW DOWIS now also contains spoken and…
2,175 results
Do What I Say (DOWIS): A Spoken Prompt Dataset for Instruction-Following NEW DOWIS now also contains spoken and…
UltraVoice: Scaling Fine-Grained Style-Controlled Speech Conversations for Spoken Dialogue Models 📝 Abstract Spoken dialogue models currently lack…
MiscSpeech-ja This dataset comprises audio and corresponding transcripts collected from a diverse range of YouTube videos and Podcasts.
Mimba PLT TTS Dataset (Plateau Malagasy Synthetic Speech)
Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English…
Risale-i Nur Sesli Külliyat — Risale-i Nur Audio Corpus
Loubna Stories Arabic Speech Dataset
NE-TTS Wancho (nnp) Cleaned TTS dataset for Wancho (nnp), a North East Indian language. Derived from the Vaani…
Emolia · Filtered · NanoCodec (FSQ) Tokens
VCTK This is a processed clone of the VCTK dataset with leading and trailing silence removed using Silero…
Speech Brain Noise Evaluation Dataset
Eka Medical ASR Evaluation Dataset
OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation…
NVSpeech Dataset Overview The NVSpeech dataset provides extensive annotations of paralinguistic vocalizations for Mandarin Chinese speech, aimed at…
Spoken-magpie LLMの日本語Instruction Tuning用データllm-jp/magpie-sft-v1.0をCosyVoice2 TTSを使用して音声化した商用利用可能な日本語の音声言語モデルのSFT用データセットです。 ある程度の話者多様性を持つように生成されています。…