Skip to content
Advertisement

Parquet

1,501 results

ImageMultimodalText

hle

!NOTE IMPORTANT: Please help us protect the integrity of this benchmark by not publicly sharing, re-uploading, or distributing…

1K–10K·MIT·Parquet
ImageMultimodalText

MMMU

This is a merged version of MMMU/MMMU with all subsets concatenated. Large-scale Multi-modality Models Evaluation Suite Accelerating the…

10K–100K·Parquet
ImageMultimodalText

DocVQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10K–100K·Apache-2.0·Parquet
Image

dtd

DTD: Describable Textures Dataset The Describable Textures Dataset (DTD) is an evolving collection of textural images in the…

1K–10K·Parquet
ImageMultimodalText

OpenFake

Dataset Card for OpenFake OpenFake is a dataset and benchmark for detecting AI-generated images, with a focus on…

1M–10M·CC-BY-NC·Parquet
ImageMultimodalText

textvqa

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10K–100K·Parquet
ImageMultimodalText

GQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10M–100M·MIT·Parquet
ImageMultimodalSensor / Time-series

libero

This dataset was created using LeRobot. Dataset Description This dataset combines four individual Libero datasets: Libero-Spatial, Libero-Object,…

100K–1M·CC-BY·Parquet
ImageMultimodalText

latex-formulas-80M

For more details, please refer to the 𝐓𝐞𝐱𝐓𝐞𝐥𝐥𝐞𝐫 GitHub repository. IMPORTANT NOTE!!! The handwritten subset of this dataset…

10M–100M·Apache-2.0·Parquet
ImageMultimodalText

geometry3k

This dataset was converted from using the following script. import json import os from datasets import Dataset, DatasetDict,…

1K–10K·MIT·Parquet
ImageMultimodalText

object365

Objects365 Dataset Objects365 detection dataset in HuggingFace parquet format. Schema Column Type Description image Image RGB image (PIL)…

100K–1M·CC-BY·Parquet
ImageMultimodalText

POPE

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10K–100K·Parquet
ImageMultimodalTabular

MapPool

MapPool - Bubbling up an extremely large corpus of maps for AI MapPool is a dataset of 75…

10M–100M·CC-BY·Parquet
Image

Food-101

Food-101

100K–1M·Custom / Research-only·Parquet
Image

Cifar100

Cifar100

10K–100K·Custom / Research-only·Parquet
ImageMultimodalText

VisionArena-Chat

VisionArena-Battle: 30K Real-World Image Conversations with Pairwise Preference Votes 200k single and multi-turn chats between users and VLM's…

100K–1M·Parquet
Advertisement