Skip to content
Advertisement

CC-BY

672 results

Audio

advanced-soundscapes-stage-1

Advanced Soundscapes Stage 1 — Raw Components (5M) This dataset contains Stage 1 output from the LAION Universal…

1M–10M·CC-BY
Audio

FSD50k

Freesound Dataset 50k (FSD50K) Important This data set is a copy from the original one located at Zenodo.…

10K–100K·CC-BY
AudioMultimodalText

IndicVoices

IndicVoices: Towards building an Inclusive Multilingual Speech Dataset for Indian Languages Updates 23 December 2025 We now have…

1M–10M·CC-BY·Parquet
AudioMultimodalText

MRSAudio

MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations Humans rely on multisensory integration to perceive…

100K–1M·CC-BY·CSV
Audio

VoxEval

VoxEval Github repository for paper: VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models Also check…

CC-BY
Advertisement