Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Summary This is the dataset proposed in our paper ICLR 2025 OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation. OpenVid-1M is a high-quality text-to-video dataset designed for research institutions to enhance video quality, featuring high aesthetics, clarity, and resolution. It can be used for direct training or as a quality tuning complement to other video datasets. All videos in the OpenVid-1M dataset have resolutions of at least 512×512.… See the full description on the dataset page:
Source: Hugging Face Hub (nkp37/OpenVid-1M). Metadata imported from the dataset’s Hub tags.