Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
The Curse of Multi-Modalities (CMM) Dataset Card Dataset details Dataset type: CMM is a curated benchmark designed to evaluate hallucination vulnerabilities in Large Multi-Modal Models (LMMs). It is constructed to rigorously test LMMs’ capabilities across visual, audio, and language modalities, focusing on hallucinations arising from inter-modality spurious correlations and uni-modal over-reliance. Dataset detail: CMM introduces 2,400 probing questions across 1… See the full description on the dataset page:
Source: Hugging Face Hub (DAMO-NLP-SG/CMM). Metadata imported from the dataset’s Hub tags.