Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Vision as Unified Multimodal Generation English 简体中文 This repository contains the dataset for the paper Vision as Unified Multimodal Generation. SenseNova Vision Corpus 50M Overview SenseNova Vision Corpus 50M (SN-VC-50M) is a large-scale multimodal vision corpus designed for unified training across diverse visual understanding and geometry-oriented tasks. The dataset is curated to address a common limitation of existing… See the full description on the dataset page:
Source: Hugging Face Hub (sensenova/SenseNova-Vision-Corpus-50M). Metadata imported from the dataset’s Hub tags.