Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Card for Symile-M3 Symile-M3 is a multilingual dataset of (audio, image, text) samples. The dataset is specifically designed to test a model’s ability to capture higher-order information between three distinct high-dimensional data types: by incorporating multiple languages, we construct a task where text and audio are both needed to predict the image, and where, importantly, neither text nor audio alone would suffice. Paper: GitHub:… See the full description on the dataset page:
Source: Hugging Face Hub (arsaporta/symile-m3). Metadata imported from the dataset’s Hub tags.