Skip to content
Advertisement

Image-to-Video

39 results

ImageMultimodalText

VideoThinkBench

CVPR 2026 Thinking with Video: Video Generation as a Promising Multimodal Reasoning Paradigm 🎊 News 2026.02 🔥🔥Our work…

1K–10K·MIT·Parquet
Image

UniVA-Bench

UniVA-Bench Paper: UniVA: Universal Video Agent towards Open-Source Next-Generation Video Generalist Project Page: Code: UniVA-Bench is a…

1K–10K·Images (folder)
3D / Point CloudImageMultimodal

DiffusionGS

ICCV 2025 DiffusionGS: Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction…

<1K·Apache-2.0
Video

openvid-wantrack-processed

OpenVid-WanTrack — preprocessed training data Preprocessed (FastVideo parquet) data used to train the TrackWan point-track-conditioned video model.…

Apache-2.0
ImageMultimodalVideo

PanFlow

PanFlow Dataset The PanFlow dataset supports the research presented in the paper PanFlow: Decoupled Motion Control for Panoramic…

CC-BY
MultimodalTextVideo

phyground

PhyGround Project Page GitHub Paper PhyGround is a criteria-grounded benchmark for evaluating physical reasoning in video generation. The…

<1K·JSON
Video

VAP-Data

Video-As-Prompt: Unified Semantic Control for Video Generation 🔥 News Oct 24, 2025: 📖 We release the first unified…

10K–100K·Apache-2.0
Advertisement