Skip to content
Advertisement

Datasets

1,204 results

Image

chemrXiv-pdf

ChemrXiv Pdf Introducing ChemrXiv Pdf, a dataset that offers access to all PDFs published until September 15, 2024.…

1K–10K·Custom / Research-only
ImageMultimodalTabular

ChEBI-20-MM

ChEBI-20-MM Dataset Overview The ChEBI-20-MM is an extensive and multi-modal benchmark developed from the ChEBI-20 dataset. It is…

10K–100K·MIT·CSV
ImageMultimodalText

Flickr30k

Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…

10K–100K·CC-BY·Parquet
ImageMultimodalText

COCO-35L

Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…

100K–1M·Parquet
ImageMultimodalText

CC3M-35L

Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…

1M–10M·Parquet
ImageMultimodalText

BEAF

BEAF: Before-After Changes for Hallucination Evaluation BEAF is a benchmark for evaluating object hallucination in vision-language models using…

10K–100K·CC-BY·Parquet
ImageMultimodalText

MultiChartQA

MultiChartQA This repository contains the questions and answers for our Multi-chart Benchmark. At present, only the data is…

1K–10K·CC-BY-NC·Parquet
Image

CARV

CARV

<1K·CC-BY·Images (folder)
ImageMultimodalText

DocVQA-2026

DocVQA 2026 ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains Building upon previous DocVQA benchmarks, this…

<1K·Parquet
Advertisement