Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
MathCanvas-Bench 🚀 Data Usage from datasets import load dataset dataset = load dataset("shiwk24/MathCanvas-Bench") print(dataset) 📖 Introduction…
MathCanvas-Bench 🚀 Data Usage from datasets import load dataset dataset = load dataset(“shiwk24/MathCanvas-Bench”) print(dataset) 📖 Introduction MathCanvas-Bench is a challenging new benchmark designed to evaluate the intrinsic Visual Chain-of-Thought (VCoT) capabilities of Large Multimodal Models (LMMs). It serves as the primary evaluation testbed for the MathCanvas framework.… See the full description on the dataset page:
Source: Hugging Face Hub (shiwk24/MathCanvas-Bench). Metadata imported from the dataset’s Hub tags.