Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
MMLU-Pro Dataset MMLU-Pro dataset is a more robust and challenging massive multi-task understanding dataset tailored to more rigorously benchmark large language models’ capabilities. This dataset contains 12K complex questions across various disciplines. Github 🏆Leaderboard đź“–Paper 🚀 What’s New 2026.03.11 Added more cutting-edge frontier models to the leaderboard, including the Claude-4.6 series, Seed2.0 series, Qwen3.5 series, and Gemini-3.1-Pro… See the full description on the dataset page:
Source: Hugging Face Hub (TIGER-Lab/MMLU-Pro). Metadata imported from the dataset’s Hub tags.