Skip to content
Advertisement

Visual Question Answering

339 results

ImageMultimodalText

synthetic-seismic-vlm

Synthetic Seismic VLM This dataset contains synthetic seismic multimodal QA rows with raw seismic images, segmentation masks, evidence-grounded…

1K–10K·MIT·Parquet
MultimodalTabularText

TerraCoT

TerraCoT TerraCoT is a remote-sensing chain-of-thought VQA dataset in which every SEG token in an answer is grounded…

1M–10M·Apache-2.0·Parquet
Image

RealText-V2

RealText-V2: A Large-Scale Multilingual Document Forgery Analysis Benchmark 💾 Dataset Description RealText-V2 is a large-scale multilingual document…

10K–100K·CC-BY-NC·Images (folder)
Image

RealText-V1

RealText-V1: A Text-Centric Image Forgery Analysis Dataset 💾 Dataset Description RealText-V1 is a text-centric image forgery analysis dataset…

1K–10K·CC-BY-NC·Images (folder)
Image

SurgRS

SurgRS

10K–100K·CC-BY-NC-SA·Images (folder)
Image

review-dataset-01

Anonymous Review Dataset This repository contains an evaluation dataset package for anonymous peer review. Overview This dataset contains…

10K–100K·CC-BY-NC·Images (folder)
Advertisement