Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Sudan-MM: A Multimodal Dataset of Sudanese Arabic Sudan-MM is the first publicly available multimodal dataset for Sudanese Arabic (السودانية), a low-resource dialect with no prior paired image-caption, video-caption, or voice-caption data. It was produced through a competitive shared task held in 2025, where five teams collected and annotated media depicting everyday Sudanese life. Each item in the dataset pairs a visual or video recording with: a written caption in Modern Standard… See the full description on the dataset page:
Source: Hugging Face Hub (IndabaXSudan/Sudan-MM). Metadata imported from the dataset’s Hub tags.