Skip to content
Advertisement

Datasets

3,216 results

ImageMultimodalText

ScienceQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10K–100K·Parquet
ImageMultimodalText

Gilt_posture_dataset

Gilt Posture Recognition Dataset Each RGB image has a matching depth image (same filename, .png extension). YOLO-format label…

1K–10K·Custom / Research-only·Images (folder)
ImageMultimodalVideo

LongVT-Source

LongVT-Source This repository contains the source video and image files for the LongVT project. Overview LongVT is an…

<1K·Apache-2.0
Image

ImageNet

ImageNet

1M–10M·Custom / Research-only·Parquet
ImageMultimodalSensor / Time-series

libero

This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v3.0", "robot type": "panda", "total episodes":…

100K–1M·Apache-2.0·Parquet
ImageMultimodalText

DSERT-RoLL

DSERT-RoLL Dataset DSERT-RoLL is a road-scene dataset collected under diverse weather and lighting conditions (e.g., Clear, Fog, Rain,…

<1K·CC-BY
ImageMultimodalText

commoncatalog-cc-by

Dataset Card for CommonCatalog CC-BY This dataset is a large collection of high-resolution Creative Common images (composed of…

10M–100M·CC-BY·Parquet
ImageMultimodalText

hot3d

HOT3D-Clips This Hugging Face repository hosts HOT3D-Clips, a set of curated sub-sequences of the HOT3D dataset. Download instructions…

100K–1M·WebDataset
Image

stanford_cars

Stanford Cars Dataset Dataset Overview Splits: Training: 8144 images used for model training. Test: 8041 images used for…

10K–100K·Parquet
ImageMultimodalText

cc3m-wds

Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated…

1M–10M·Custom / Research-only·WebDataset
Image

gtsrb

Dataset Card for German Traffic Sign Recognition Benchmark This dataset contains images of 43 classes of traffic signs.…

100K–1M·Parquet
ImageMultimodalVideo

360Motion-Dataset

360°-Motion Dataset Project page Paper Code Acknowledgments We thank Jinwen Cao, Yisong Guo, Haowen Ji, Jichao Wang, and…

<1K·Apache-2.0·Images (folder)
ImageMultimodalText

fine-t2i

Fine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning arxiv by Xu Ma, Yitian Zhang, Qihua…

100K–1M·Apache-2.0·WebDataset
Image

OmniDocBench

OmniDocBench English 简体中文 OmniDocBench is an evaluation dataset for diverse document parsing in real-world scenarios, with the following…

1K–10K·Images (folder)
Advertisement