Skip to content
Advertisement

CC-BY-NC

210 results

Text

BioMatrix-SFT

BioMatrix-SFT This is the supervised fine-tuning (SFT) / instruction-tuning corpus used to train BioMatrix, a multimodal foundation model…

10M–100M·CC-BY-NC·Parquet
ImageMultimodalText

MultiChartQA

MultiChartQA This repository contains the questions and answers for our Multi-chart Benchmark. At present, only the data is…

1K–10K·CC-BY-NC·Parquet
ImageMultimodalText

S1-MMAlign

S1-MMAlign A Large-Scale Multi-Disciplinary Scientific Multimodal Dataset S1-MMAlign is a large-scale, multi-disciplinary multimodal dataset…

10M–100M·CC-BY-NC·WebDataset
ImageMultimodalText

SDG-30K

SDG-30K — Structured Defect Grounding Dataset A 30,000-image dataset for structured defect grounding in text-to-image generations. Each image…

10K–100K·CC-BY-NC·JSON
Image

VLM-SubtleBench

VLM-SubtleBench VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning? The ability to distinguish subtle differences…

10K–100K·CC-BY-NC
ImageMultimodalText

Zebra-CoT

Zebra‑CoT A diverse large-scale dataset for interleaved vision‑language reasoning traces. Dataset Description Zebra‑CoT is a diverse large‑scale…

100K–1M·CC-BY-NC·Parquet
Image

WorldMemArena

WorldMemArena WorldMemArena is a large-scale multimodal memory benchmark designed to evaluate how well AI systems retain, update, and…

10K–100K·CC-BY-NC·Images (folder)
Advertisement