Skip to content
Advertisement

Video

706 results

3D / Point CloudImageMultimodal

DiffusionGS

ICCV 2025 DiffusionGS: Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction…

<1K·Apache-2.0
MultimodalTabularText

MSR-VTT

Clone from "friedrichor/MSR-VTT". MSRVTT contains 10K video clips and 200K captions. We adopt the standard 1K-A split protocol,…

10K–100K·JSON
MultimodalTabularText

MVTamperBench

MVTamperBench Dataset Overview MVTamperBench is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT
MultimodalTabularVideo

deas_robocasa

DEAS-RoboCasa Robocasa dataset used for fine-tuning GR00T-N1.5 in DEAS Dataset Owner(s): Changyeon Kim Dataset Creation Date: October 14,…

100K–1M·MIT·Parquet
MultimodalSensor / Time-seriesTabular

so-combined-ru

Датасет создан при помощи библиотеки LeRobot. Описание датасета Русскоязычная версия данного датасета объединяет 598 открытых датасетов сообщества в…

1M–10M·Apache-2.0·Parquet
Advertisement