openwhisper
Streaming ASR Dataset This dataset is designed for training real-time (streaming) ASR models, with a focus on handling…
2,147 results
Streaming ASR Dataset This dataset is designed for training real-time (streaming) ASR models, with a focus on handling…
IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages Updates 23 December 2025 We now have…
Crowd Whatsapp Yiddish - Source Dataset
Knesset Plenum Source Dataset
Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream…
ESB Test Sets: Parquet & Sorted This dataset takes the open-asr-leaderboard/datasets-test-only data and sorts each split by audio…
FreeSound.org LAION-640k Dataset
JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation HomePage Paper GitHub TL;DR We introduce JavisGPT, a…
AVQA JSONL (Audio Multiple-Choice QA)
NatureLM-audio-training
Synthetic ASR data — zh Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream…
Malaysian Youtube Malaysian and Singaporean youtube channels, total up to 60k audio files with total 18.7k hours. URLs…