Skip to content
Advertisement

CC-BY

672 results

ImageMultimodalText

GeoDrive-Bench

GeoDrive-Bench A multi-country driving scene benchmark for evaluating vision-language models on culture- and region-specific traffic knowledge.…

1K–10K·CC-BY·Parquet
ImageMultimodalSensor / Time-series

libero

This dataset was created using LeRobot. Dataset Description This dataset combines four individual Libero datasets: Libero-Spatial, Libero-Object,…

100K–1M·CC-BY·Parquet
ImageMultimodalText

object365

Objects365 Dataset Objects365 detection dataset in HuggingFace parquet format. Schema Column Type Description image Image RGB image (PIL)…

100K–1M·CC-BY·Parquet
ImageMultimodalTabular

MapPool

MapPool - Bubbling up an extremely large corpus of maps for AI MapPool is a dataset of 75…

10M–100M·CC-BY·Parquet
Image

dave_sonar

OA-Stereo Opti-Acoustic Underwater Stereo Dataset Overview This dataset accompanies the paper: OA-Stereo: Self-Supervised Opti-Acoustic Stereo for…

10K–100K·CC-BY·Images (folder)
Image

datacomp_pools

DataComp Pools This repository contains metadata files for DataComp. For details on how to use the metadata, please…

CC-BY
Image

COCO

330K images with object, segmentation and caption labels.

100K–1M·CC-BY·COCO JSON
Audio

AudioSet

2M+ human-labeled 10s clips across 600+ classes.

1M–10M·CC-BY·Audio (wav/mp3)
Audio

LibriSpeech

~1,000 hours of aligned read English speech.

100K–1M·CC-BY·Audio (wav/mp3)
Advertisement