Skip to content
Advertisement

Open

2,922 results

ImageMultimodalText

lines-dataset

P&ID Line Detection Dataset This dataset contains cropped images from P&ID (Piping and Instrumentation Diagrams) with line segment…

10K–100K·MIT·Images (folder)
ImageMultimodalTabular

GUI-Odyssey

Dataset Card for GUI Odyssey News⭐️ A new and improved version of the GUIOdyssey dataset has been released!…

1K–10K·CC-BY·JSON
ImageMultimodalText

ChartQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

1K–10K·Parquet
ImageMultimodalText

Jensky

baseball-detection-2 2023-06-02 3:09pm Provided by a Roboflow user License: CC BY 4.0 baseball-detection-2 - v4 2023-06-02 3:09pm This…

1K–10K·MIT·Images (folder)
ImageMultimodalText

ScienceQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…

10K–100K·Parquet
ImageMultimodalText

Gilt_posture_dataset

Gilt Posture Recognition Dataset Each RGB image has a matching depth image (same filename, .png extension). YOLO-format label…

1K–10K·Custom / Research-only·Images (folder)
ImageMultimodalVideo

LongVT-Source

LongVT-Source This repository contains the source video and image files for the LongVT project. Overview LongVT is an…

<1K·Apache-2.0
Image

ImageNet

ImageNet

1M–10M·Custom / Research-only·Parquet
ImageMultimodalSensor / Time-series

libero

This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v3.0", "robot type": "panda", "total episodes":…

100K–1M·Apache-2.0·Parquet
ImageMultimodalText

DSERT-RoLL

DSERT-RoLL Dataset DSERT-RoLL is a road-scene dataset collected under diverse weather and lighting conditions (e.g., Clear, Fog, Rain,…

<1K·CC-BY
ImageMultimodalText

commoncatalog-cc-by

Dataset Card for CommonCatalog CC-BY This dataset is a large collection of high-resolution Creative Common images (composed of…

10M–100M·CC-BY·Parquet
ImageMultimodalText

hot3d

HOT3D-Clips This Hugging Face repository hosts HOT3D-Clips, a set of curated sub-sequences of the HOT3D dataset. Download instructions…

100K–1M·WebDataset
Image

stanford_cars

Stanford Cars Dataset Dataset Overview Splits: Training: 8144 images used for model training. Test: 8041 images used for…

10K–100K·Parquet
ImageMultimodalText

cc3m-wds

Dataset Card for Conceptual Captions (CC3M) Dataset Summary Conceptual Captions is a dataset consisting of ~3.3M images annotated…

1M–10M·Custom / Research-only·WebDataset
Image

gtsrb

Dataset Card for German Traffic Sign Recognition Benchmark This dataset contains images of 43 classes of traffic signs.…

100K–1M·Parquet
Advertisement