Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Visualize on Visual Layer Imagenet-1K-VL-Enriched An enriched version of the ImageNet-1K Dataset with image caption, bounding boxes, and label…
Visualize on Visual Layer Imagenet-1K-VL-Enriched An enriched version of the ImageNet-1K Dataset with image caption, bounding boxes, and label issues! With this additional information, the ImageNet-1K dataset can be extended to various tasks such as image retrieval or visual question answering. The label issues helps to curate a cleaner and leaner dataset. Description The dataset consists of 6 columns: image id: The original filename of the image from… See the full description on the dataset page:
Source: Hugging Face Hub (visual-layer/imagenet-1k-vl-enriched). Metadata imported from the dataset’s Hub tags.