Skip to content
Advertisement
ImageMultimodalText

hle

!NOTE IMPORTANT: Please help us protect the integrity of this benchmark by not publicly sharing, re-uploading, or distributing the dataset.…

!NOTE IMPORTANT: Please help us protect the integrity of this benchmark by not publicly sharing, re-uploading, or distributing the dataset. Humanity’s Last Exam 🌐 Website 📄 Paper GitHub Center for AI Safety & Scale AI Humanity’s Last Exam (HLE) is a multi-modal benchmark at the frontier of human knowledge, designed to be the final closed-ended academic benchmark of its kind with broad subject coverage. Humanity’s Last Exam consists of 2,500 questions across dozens… See the full description on the dataset page:

Source: Hugging Face Hub (cais/hle). Metadata imported from the dataset’s Hub tags.

Advertisement