Open
2,922 results
dowis
Do What I Say (DOWIS): A Spoken Prompt Dataset for Instruction-Following NEW DOWIS now also contains spoken and…
UltraVoice
UltraVoice: Scaling Fine-Grained Style-Controlled Speech Conversations for Spoken Dialogue Models 📝 Abstract Spoken dialogue models currently lack…
MiscSpeech-ja
MiscSpeech-ja This dataset comprises audio and corresponding transcripts collected from a diverse range of YouTube videos and Podcasts.
turkish-tts-combined-raw
Türkçe TTS Birleşik Veri Seti 7 farklı açık kaynak Türkçe TTS veri setinin birleşimi. ~81,500 örnek 24kHz SNAC…
mls-annotated
Dataset Card for Annotations of non English MLS This dataset consists in annotations of a the Non English…
Risale-i Nur Sesli Külliyat — Risale-i Nur Audio Corpus
Risale-i Nur Sesli Külliyat — Risale-i Nur Audio Corpus
Loubna Stories Arabic Speech Dataset
Loubna Stories Arabic Speech Dataset
ne-tts-nnp
NE-TTS Wancho (nnp) Cleaned TTS dataset for Wancho (nnp), a North East Indian language. Derived from the Vaani…
Emolia · Filtered · NanoCodec (FSQ) Tokens
Emolia · Filtered · NanoCodec (FSQ) Tokens
vctk
VCTK This is a processed clone of the VCTK dataset with leading and trailing silence removed using Silero…
Speech Brain Noise Evaluation Dataset
Speech Brain Noise Evaluation Dataset
Eka Medical ASR Evaluation Dataset
Eka Medical ASR Evaluation Dataset
OpenDialog
OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation…
hawrami-kurdish-raw-audio
Hawrami Raw Audio Collection Overview This repository contains approximately 500 hours of Hawrami Kurdish raw speech collected from…