Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Summary This is the dataset proposed in our paper VidProM: A Million-scale Real Prompt-Gallery Dataset for Text-to-Video Diffusion Models (NeurIPS 2024). VidProM is the first dataset featuring 1.67 million unique text-to-video prompts and 6.69 million videos generated from 4 different state-of-the-art diffusion models. It inspires many exciting new research areas, such as Text-to-Video Prompt Engineering, Efficient Video Generation, Fake Video Detection, and Video Copy… See the full description on the dataset page:
Source: Hugging Face Hub (WenhaoWang/VidProM). Metadata imported from the dataset’s Hub tags.