Skip to content
Advertisement

Text

2,175 results

MultimodalTabularText

Polymarket_data

Polymarket Data Complete Data Infrastructure for Polymarket — Fetch, Process, Analyze A comprehensive dataset of 1.9 billion trading…

>1B
MultimodalTabularText

molmobot-data

MolmoBot-data Training episode data (actions, visual inputs, and other sensor data) for 8 tasks on 2 robotic platforms:…

100K–1M·ODC-BY·Parquet
MultimodalTextVideo

VSI-Bench

Dataset arXiv Website Code VSI-Bench VSI-Bench-Debiased !IMPORTANT Nov. 7, 2025 UPDATE: This Dataset has been updated to include…

10K–100K·Apache-2.0·Parquet
MultimodalTextVideo

M3arsSynth

Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions M3arsSynth Dataset Summary M3arsSynth is a large-scale,…

CC-BY
ImageMultimodalText

RoboCerebra

Overview RoboCerebra is a large-scale benchmark designed to evaluate robotic manipulation in long-horizon tasks, shifting the focus of…

1K–10K·MIT·Parquet
MultimodalTextVideo

phyworldbench

PhyWorldBench This repository hosts the core assets of PhyWorldBench, the 1,050 JSON prompt files, the evaluation standards, and…

<1K·MIT·Parquet
MultimodalTextVideo

REVISOR-25k

REVISOR-25k A multi-task video understanding dataset for training video LLMs with reinforcement learning (GRPO). The dataset contains ~25k…

10K–100K·Apache-2.0·Parquet
MultimodalTextVideo

mcd_rppg

MCD-rPPG: Multi-Camera Dataset for Remote Photoplethysmography This repository contains the dataset from the paper "Gaze into the Heart:…

1K–10K·CC-BY
MultimodalTextVideo

CultureScore

CULTURESCORE: Evaluating Cultural Faithfulness in Video Generation Models Dataset Summary CultureScore is a benchmark dataset for evaluating cultural…

1K–10K·CC-BY
ImageMultimodalTabular

MentalBlackboard

MentalBlackboard Benchmark This repository contains the MentalBlackboard benchmark with multiple tasks: - Prediction - Planning Dataset Sources

1K–10K·CC-BY·Parquet
ImageMultimodalText

AVGen-Bench

AVGen-Bench Generated Videos Data Card Overview This data card describes the generated audio-video outputs stored directly in the…

1K–10K·MIT·Parquet
MultimodalTextVideo

fMRI-Shape

fMRI-Shape Dataset: A Component of the fMRI-3D Dataset for MinD-3D++ This repository contains the fMRI-Shape dataset, a component…

1K–10K·Apache-2.0·Text (raw)
MultimodalTextVideo

VideoChat2

Video training data of LongVU downloaded from Video Please download the original videos from the provided links: BDD100K:…

100K–1M·MIT·JSON
MultimodalTextVideo

RoboReward

RoboReward Links: Paper · RoboRewardBench Leaderboard RoboReward is a dataset for training and evaluating general-purpose vision-language reward…

10K–100K·CC-BY
Advertisement