Skip to content
Advertisement

Datasets

634 results

3D / Point CloudImageMultimodal

DiffusionGS

ICCV 2025 DiffusionGS: Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction…

<1K·Apache-2.0
MultimodalTabularText

MSR-VTT

Clone from "friedrichor/MSR-VTT". MSRVTT contains 10K video clips and 200K captions. We adopt the standard 1K-A split protocol,…

10K–100K·JSON
MultimodalTabularText

MVTamperBench

MVTamperBench Dataset Overview MVTamperBench is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT
Advertisement