Skip to content
Advertisement

Multiple Choice

46 results

MultimodalTabularText

MapEval-API

MapEval-API MapEval-API is created using MapQaTor. Usage from datasets import load dataset Load dataset ds = load dataset("MapEval/MapEval-API",…

<1K·Apache-2.0·JSON
ImageMultimodalText

MapEval-Visual

MapEval-Visual This dataset was introduced in MapEval: A Map-Based Evaluation of Geo-Spatial Reasoning in Foundation Models Example Query…

<1K·Apache-2.0·JSON
ImageMultimodalText

CHOICE

CHOICE: Benchmarking The Remote Sensing Capabilities of Large Vision-Language Models Abstract: The rapid advancement of Large Vision-Language Models…

1K–10K·MIT·Images (folder)
ImageMultimodalText

PhyX

PhyX: Does Your Model Have the "Wits" for Physical Reasoning? Dataset for the paper "PhyX: Does Your Model…

10K–100K·MIT·Parquet
ImageMultimodalText

MathVerse

Dataset Card for MathVerse Dataset Description Paper Information Dataset Examples Leaderboard Citation Dataset Description The capabilities of…

1K–10K·MIT·Parquet
ImageMultimodalText

ship-dataset

ShipBench: A Drawing-Grounded VLM Benchmark for Ship Structural Reasoning ShipBench is a metadata-grounded vision-language benchmark on…

10K–100K·CC-BY·JSON
ImageMultimodalText

XLRS-Bench-lite

🐙GitHub Information or evaluatation on this dataset can be found in this repo: 📜Dataset License Annotations of this…

1K–10K·CC-BY-NC-SA·Arrow
Advertisement