Skip to content
Advertisement

Open

2,922 results

Text

OPUS-100

OPUS-100

10M–100M·Custom / Research-only·Parquet
Text

WMT14

WMT14

10M–100M·Custom / Research-only·Parquet
Text

OpusBooks

OpusBooks

1M–10M·Custom / Research-only·Parquet
Text

PHINC

Abstract Code-mixing is the phenomenon of using more than one language in a sentence. In the multilingual communities,…

10K–100K·CC-BY·CSV
Text

multi30k

Multi30k This dataset contains the "multi30k" dataset, which is the "task 1" dataset from here. Each example consists…

10K–100K·JSON
ImageMultimodalText

BEAF

BEAF: Before-After Changes for Hallucination Evaluation BEAF is a benchmark for evaluating object hallucination in vision-language models using…

10K–100K·CC-BY·Parquet
ImageMultimodalText

MultiChartQA

MultiChartQA This repository contains the questions and answers for our Multi-chart Benchmark. At present, only the data is…

1K–10K·CC-BY-NC·Parquet
Image

CARV

CARV

<1K·CC-BY·Images (folder)
ImageMultimodalText

DocVQA-2026

DocVQA 2026 ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains Building upon previous DocVQA benchmarks, this…

<1K·Parquet
ImageMultimodalText

GQA-ru

GQA-ru This is a translated version of original GQA dataset and stored in format supported for lmms-eval pipeline.…

10K–100K·Apache-2.0·Parquet
Advertisement