Skip to content
Advertisement

MIT

552 results

MultimodalTabularText

MVTamperBench

MVTamperBench Dataset Overview MVTamperBench is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT
MultimodalTabularVideo

deas_robocasa

DEAS-RoboCasa Robocasa dataset used for fine-tuning GR00T-N1.5 in DEAS Dataset Owner(s): Changyeon Kim Dataset Creation Date: October 14,…

100K–1M·MIT·Parquet
ImageMultimodalTabular

MVBench

MVBench Important Update 18/10/2024 Due to NTU RGB+D License, 320 videos from NTU RGB+D need to be downloaded…

1K–10K·MIT·JSON
MultimodalTabularText

uci-har-federated

UCI Human Activity Recognition (HAR) Dataset Dataset Description The UCI Human Activity Recognition dataset is a widely-used benchmark…

10K–100K·MIT·Parquet
MultimodalTabularText

pg-en

Overview Property Value Source Project Gutenberg (English catalog) Snapshot 2026-07-02-18-47-04 Total files 50871 Total Tokens (BPE) ~7.14 billion…

10K–100K·MIT·Parquet
MultimodalTabularText

binance-btcusdt

BTCUSDT Perpetual Futures — 5-Minute Feature Dataset Complete historical dataset for Binance BTCUSDT USDT-Margined Perpetual Futures, covering…

>1B·MIT·Parquet
MultimodalTabularText

TraitGym

🧬 TraitGym Benchmarking DNA Sequence Models for Causal Regulatory Variant Prediction in Human Genetics 🏆 Leaderboard: ⚡️ Quick…

10M–100M·MIT·Parquet
MultimodalTabularText

mrcr

OpenAI MRCR: Long context multiple needle in a haystack benchmark OpenAI MRCR (Multi-round co-reference resolution) is a long…

1K–10K·MIT·Parquet
MultimodalTabularText

MMLU-ProX

MMLU-ProX MMLU-ProX is a multilingual benchmark that builds upon MMLU-Pro, extending to 29 typologically diverse languages, designed to…

100K–1M·MIT·Parquet
MultimodalTabularText

betty-dota2

Betty Dota 2 — Decision Context Dataset Overview 9,385 professional Dota 2 matches parsed from replay files (.dem)…

>1B·MIT·Parquet
Advertisement