A multi-hazard, multi-sensor, and multi-task vision-language dataset for global-scale disaster asses
A multi-hazard, multi-sensor, and multi-task vision-language dataset for global-scale disaster assessment and response.
122 results
A multi-hazard, multi-sensor, and multi-task vision-language dataset for global-scale disaster assessment and response.
UAVReason VQA / Caption / Generation
TriSearch-v1 (0.0.1)
A multi-hazard, multi-sensor, and multi-task vision-language dataset for global-scale disaster assessment and response.
TAMMs: Change Understanding and Forecasting in Satellite Image Time Series with a Temporal-Aware Multimodal Model 📄 Paper (ICLR…
MMS-VPR: A Fine-Grained Multimodal Street-Level Visual Place Recognition Dataset and Evaluation Benchmark for Dense Pedestrian Environments
LUCID – Lunar Captioned Image Dataset
Blackline Atlas Training Corpus v1 Blackline Atlas Training Corpus v1 is a license-aware, normalized dataset for building structured…
Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…
LAION-COCO translated to 200 languages
Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…
Marathi OCR model training dataset
Team and Homepage Official Website: Hugging Face Organization: Contact If you encounter any issues with the dataset or…
S1-MMAlign A Large-Scale Multi-Disciplinary Scientific Multimodal Dataset S1-MMAlign is a large-scale, multi-disciplinary multimodal dataset…
ICDAR2019's Scanned Receipts OCR and Information Extraction (SROIE)
QCalEval Dataset Description The dataset contains scientific plots from quantum computing calibration experiments, paired with vision-language…