Skip to content
Advertisement

Gated

293 results

Video

Galaxea-Open-World-Dataset

Galaxea Open-World Dataset Key Features 500+ hours of real-world mobile manipulation data. All data collected using one uniform…

>1B·CC-BY-NC-SA
AudioMultimodalText

mu-bench

Dataset Card for μ-Bench (Leaderboard Code) μ-Bench is a multilingual transcription benchmark built from real customer-service phone conversations.…

1K–10K·CC-BY-NC·Audio (folder)
Audio

DailyTalkContiguous-MoodyGirl

Moody Girl: Emotion-Tagged DailyTalk Dataset This dataset is an emotion-tagged version of the DailyTalk dataset, enhanced with emotion…

CC-BY-SA
AudioMultimodalText

commonvoice22_sidon

CV22-Sidon Overview This dataset hosts a release of Mozilla Common Voice 22 restored with the Sidon speech restoration…

10M–100M·CC0·WebDataset
AudioMultimodalText

openwhisper

Streaming ASR Dataset This dataset is designed for training real-time (streaming) ASR models, with a focus on handling…

100K–1M·MIT·Parquet
AudioMultimodalText

IndicVoices

IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages Updates 23 December 2025 We now have…

1M–10M·CC-BY·Parquet
AudioMultimodalText

synthetic-asr-hi

Synthetic ASR data — hi Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream…

10K–100K·JSON
Audio

finetuned-hindi-punjabi-denoised

Multilingual Speaker Diarization Dataset This dataset contains synthetic multilingual speaker diarization data with Hindi, English, and Punjabi audio…

1K–10K·MIT
AudioMultimodalText

synthetic-asr-zh

Synthetic ASR data — zh Generated by Valsea-ASR/synthetic-data-pipeline. Audio is synthetic (TTS), targeted as training data for downstream…

10K–100K·JSON
Advertisement