Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Llama3-SSL4EO-S12-Captions The captions are aligned with the SSL4EO-S12 v1.1 dataset and were automatically generated using the Llama3-LLaVA-Next-8B model. Please find more information regarding the generation and evaluation in the Llama3-MS-CLIP paper. Code: Data Structure We provide the captions in two versions: As a single compressed Parquet file per split and as CSV files with 256 captions each that match the Zarr Zip files of the… See the full description on the dataset page:
Source: Hugging Face Hub (ibm-esa-geospatial/Llama3-SSL4EO-S12-v1.1-captions). Metadata imported from the dataset’s Hub tags.