Skip to content
Advertisement

English

1,233 results

MultimodalTextVideo

REVISOR-25k

REVISOR-25k A multi-task video understanding dataset for training video LLMs with reinforcement learning (GRPO). The dataset contains ~25k…

10K–100K·Apache-2.0·Parquet
MultimodalTextVideo

CultureScore

CULTURESCORE: Evaluating Cultural Faithfulness in Video Generation Models Dataset Summary CultureScore is a benchmark dataset for evaluating cultural…

1K–10K·CC-BY
ImageMultimodalTabular

MentalBlackboard

MentalBlackboard Benchmark This repository contains the MentalBlackboard benchmark with multiple tasks: - Prediction - Planning Dataset Sources

1K–10K·CC-BY·Parquet
ImageMultimodalVideo

LoVoRA

LoVoRA Dataset: Text-guided and Mask-free Video Object Removal and Addition Authors: Zhihan Xiao, Lin Liu, Yixin Gao, Xiaopeng…

10K–100K·Apache-2.0
Video

robotwin-icl-paired-v3

RoboTwin ICL-paired v3 Cross-embodiment paired dataset on RoboTwin: 25 manipulation tasks rendered for 3 robots (arx-x5, franka, ur5)…

10K–100K·Apache-2.0
ImageMultimodalVideo

PanFlow

PanFlow Dataset The PanFlow dataset supports the research presented in the paper PanFlow: Decoupled Motion Control for Panoramic…

CC-BY
Video

VAP-Data

Video-As-Prompt: Unified Semantic Control for Video Generation 🔥 News Oct 24, 2025: 📖 We release the first unified…

10K–100K·Apache-2.0
Video

AgiBot-World-Beta-no-torso-movement

Dataset taken from BAAI-DataCube/agibot-lerobot-v30-217-task-collection, in LeRobot Datasets v3.0 format. Fisheye video features are dropped so there…

10K–100K·CC-BY-NC-SA
Video

PhysicTran38K

PhysicTran38K We provide 38K video-based dataset for physics-aware image editing, by casting editing as physical state transitions. To…

10K–100K·Apache-2.0
Advertisement