Skip to content
Advertisement

CC-BY-NC

210 results

ImageMultimodalVideo

GOKU-2M

Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editing GOKU-2M is a large-scale, unified instruction-based…

1M–10M·CC-BY-NC
MultimodalTextVideo

CSL-News

Summary This is the dataset proposed in our paper "Uni-Sign: Toward Unified Sign Language Understanding at Scale". CSL-News…

100K–1M·CC-BY-NC·JSON
Audio

LibriBrain

LibriBrain (Sherlock Holmes 1–7) Paper Code This repository contains the LibriBrain data organised by book: MEG recordings (.h5),…

CC-BY-NC
AudioMultimodalText

urbansound8K

(card and dataset copied from This dataset contains 8732 labeled sound excerpts (<=4s) of urban sounds from 10…

1K–10K·CC-BY-NC·Parquet
AudioMultimodalText

mu-bench

Dataset Card for μ-Bench (Leaderboard Code) μ-Bench is a multilingual transcription benchmark built from real customer-service phone conversations.…

1K–10K·CC-BY-NC·Audio (folder)
AudioMultimodalTabular

multi_lingo_data

Multi-lingual TTS Data (leeoxiang/multi lingo data) Large-scale cross-lingual + same-lingual TTS corpus for training expressive multi-lingual…

1K–10K·CC-BY-NC
Audio

WearVox

WearVox: An Egocentric Multichannel Voice Assistant Benchmark for Wearables Paper: WearVox: An Egocentric Multichannel Voice Assistant Benchmark for…

1K–10K·CC-BY-NC·Audio (folder)
Audio

AIR-Bench-Dataset

AIR-Bench Arxiv: is the AIR-Bench dataset download page.AIR-Bench encompasses two dimensions: foundation and chat benchmarks. The former consists…

<1K·CC-BY-NC·Audio (folder)
Advertisement