Skip to content
Advertisement

English

1,233 results

MultimodalTabularText

osm-polygon-selection

osm-polygon-selection dataset A curated set of OpenStreetMap polygons from 310 geographic units — sovereign countries plus sub-country regions…

10M–100M·ODbL·Parquet
MultimodalTabularText

medicine-tasks

Adapting LLMs to Domains via Continual Pre-Training (ICLR 2024) This repo contains the evaluation datasets for our paper…

1K–10K·JSON
MultimodalTabularText

reward-bench-2

Code Leaderboard Results Paper RewardBench 2 Evaluation Dataset Card The RewardBench 2 evaluation dataset is the new version…

1K–10K·ODC-BY·Parquet
MultimodalTabularText

MVTamperBench

MVTamperBench Dataset Overview MVTamperBench is a robust benchmark designed to evaluate Vision-Language Models (VLMs) against adversarial video…

10K–100K·MIT
ImageMultimodalTabular

MVBench

MVBench Important Update 18/10/2024 Due to NTU RGB+D License, 320 videos from NTU RGB+D need to be downloaded…

1K–10K·MIT·JSON
MultimodalTabularText

agent-llm-traces

Multi-Benchmark LLM Agent Traces A comprehensive dataset of OpenTelemetry traces capturing LLM inference behavior across multiple agent frameworks,…

1K–10K·Other·Parquet
MultimodalTabularText

KodCode-V1

🐱 KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding KodCode is the largest fully-synthetic open-source dataset…

100K–1M·CC-BY-NC·Parquet
MultimodalTabularText

binance-btcusdt

BTCUSDT Perpetual Futures — 5-Minute Feature Dataset Complete historical dataset for Binance BTCUSDT USDT-Margined Perpetual Futures, covering…

>1B·MIT·Parquet
Advertisement