Multilingual
575 results
Emolia-Thinking (VoiceNet balanced subset)
Emolia-Thinking (VoiceNet balanced subset)
nonverbalspeech38k
🎉 🎉 🎉 NonVerbalSpeech-38K: A Scalable Pipeline for Enabling Non-Verbal Speech Generation and Understanding The official repository for…
SimbaBench_dataset
SibmaBench Data Release & Benchmarking To evaluate your model on SimbaBench across all supported tasks (ASR, TTS, and…
SpeechJudge-Data
SpeechJudge-Data: A Large-Scale Human Feedback Corpus for Speech Generation Introduction SpeechJudge-Data is a large-scale human feedback corpus of…
Open Bible Resources — African Languages
Open Bible Resources — African Languages
moss-character-voices-bestof64
MOSS Character Voices — Best-of-64 (Stage 2) Best-of-64 voice-acting takes from the 4.55B MOSS-TTS-Local voice-acting model…
emolia-thinking-balanced-buckets
Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that…
Codemixed_New
Codemixed ASR Dataset Unified collection of code-mixed ASR datasets.
INTP
INTP: Intelligibility Preference Speech Dataset We establish a synthetic Intelligibility Preference Speech Dataset (INTP), including about 250K…