MMStar
MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage 🤗 Dataset 🤗 Paper…
1,204 results
MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage 🤗 Dataset 🤗 Paper…
WebLINX: Real-World Website Navigation with Multi-Turn Dialogue WARNING: This is not the main WebLINX data card! You might…
Chitralekha Dataset Details Dataset Version Some of the fonts do not have proper letters/rendering of different telugu letter…
FishingROV: scallop lr teacher 640 This dataset is an optimized derivative format generated for the FishingROV edge inference…
TN5000 Thyroid Nodule Classification (Cropped, 224×224)
Dataset Card for Industry Documents Library (IDL) Dataset Summary Industry Documents Library (IDL) is a document dataset filtered…
Cityscape‑Adverse A benchmark for evaluating semantic segmentation robustness under realistic adverse conditions. Overview Cityscape‑Adverse extends…
A multi-hazard, multi-sensor, and multi-task vision-language dataset for global-scale disaster assessment and response.
QIN-LungCT-Seg (multi-site lung nodule segmentations)
Synthetic Living Room Dataset for Robotic Perception Generated by datapack-import.ts This dataset mirrors public data-pack render outputs from…
RealText-V2: A Large-Scale Multilingual Document Forgery Analysis Benchmark 💾 Dataset Description RealText-V2 is a large-scale multilingual document…
LayeredFlow-Syn Extracted Ground Truth
Chest X-Ray Restoration & Classification — Dataset + Pipeline Outputs Companion dataset for the project Restorasi dan Klasifikasi…