Skip to content
Advertisement

10K–100K

748 results

Video

memo_data

MEMO Video Dataset This dataset contains a curated collection of human talking videos gathered from publicly accessible sources…

10K–100K·CC-BY
Video

siftformer2-data

SiftFormer2 Data K400 / SSv2 video classification 학습용 데이터. 구성 siftformer2-data/ ├── k400/ │ ├── train manifest.jsonl │…

10K–100K·MIT
MultimodalTextVideo

EyePCR

A dataset for "EyePCR: A Comprehensive Benchmark for Fine-Grained Perception, Knowledge Comprehension and Clinical Reasoning in Ophthalmic Surgery"…

10K–100K·CC-BY·JSON
Video

Vera-Layered-Video-Dataset

Dataset for Vera: A Layered Diffusion Model for Content-Preserving Video Editing Hongkai Zheng¹²  ·  Ta-Ying Cheng²  ·  Benjamin…

10K–100K·Apache-2.0
Video

FineGym-skeleton

FineGym-skeleton Dataset License: CC-BY-4.0 Overview The FineGym-skeleton dataset is a human skeleton-based action recognition benchmark derived from…

10K–100K·CC-BY
Video

PusaV0.5_Training

PusaV0.5 Training Dataset Code Repository Model Hub Training Toolkit Dataset Pusa Paper FVDM Paper Follow on X Xiaohongshu…

10K–100K·Apache-2.0
Video

MAVOS-DD

LICENSE: This dataset is released under the CC BY-NC-SA 4.0 license. This repository contains MAVOS-DD an open-set benchmark…

10K–100K
Video

EgoLife

Data cleaning, stay tuned! Please refer to first for general info. Checkout the paper EgoLife ( for more…

10K–100K·MIT
Video

MAVOS-DD

LICENSE: This dataset is released under the CC BY-NC-SA 4.0 license. This repository contains MAVOS-DD an open-set benchmark…

10K–100K
Video

MAVOS-DD

LICENSE: This dataset is released under the CC BY-NC-SA 4.0 license. This repository contains MAVOS-DD an open-set benchmark…

10K–100K
AudioMultimodalText

meow-10k

Dataset Card for Meow-10K Meow-10K is a high-fidelity, synchronized quad-modal dataset comprising 10,000 feline samples. It is the…

10K–100K·Apache-2.0·JSON
Advertisement