Skip to content
Advertisement

Visual Question Answering

339 results

ImageMultimodalText

MapEval-Visual

MapEval-Visual This dataset was introduced in MapEval: A Map-Based Evaluation of Geo-Spatial Reasoning in Foundation Models Example Query…

<1K·Apache-2.0·JSON
Image

UltraVR

UltraVR

<1K·Custom / Research-only·Images (folder)
ImageMultimodalText

HRVQA-2k

A 2k subset of the validation split of the HRVQA dataset ported to HF for ease-of-use in quick…

1K–10K·CC-BY-NC·Parquet
ImageMultimodalText

RSVQA-LR-2k

A 2k subset of the validation split of the RSVQA LR dataset ported to HF for ease-of-use in…

1K–10K·CC-BY·Parquet
ImageMultimodalText

RSVQA-HR-2k

A 2k subset of the validation split of the RSVQA HR dataset ported to HF for ease-of-use in…

1K–10K·CC-BY·Parquet
Image

blackline-atlas-training-corpus-v1

Blackline Atlas Training Corpus v1 Blackline Atlas Training Corpus v1 is a license-aware, normalized dataset for building structured…

<1K·Custom / Research-only·Images (folder)
ImageMultimodalText

BEAF

BEAF: Before-After Changes for Hallucination Evaluation BEAF is a benchmark for evaluating object hallucination in vision-language models using…

10K–100K·CC-BY·Parquet
ImageMultimodalText

MultiChartQA

MultiChartQA This repository contains the questions and answers for our Multi-chart Benchmark. At present, only the data is…

1K–10K·CC-BY-NC·Parquet
Image

CARV

CARV

<1K·CC-BY·Images (folder)
ImageMultimodalText

DocVQA-2026

DocVQA 2026 ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains Building upon previous DocVQA benchmarks, this…

<1K·Parquet
Advertisement