Parquet
1,501 results
synthetic-living-room-dataset-for-robotic-perception
Synthetic Living Room Dataset for Robotic Perception Generated by datapack-import.ts This dataset mirrors public data-pack render outputs from…
LayeredFlow-Syn Extracted Ground Truth
LayeredFlow-Syn Extracted Ground Truth
SEED-Bench
Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…
oxford-iiit-pet
The Oxford-IIIT Pet Dataset Description A 37 category pet dataset with roughly 200 images for each class. The…
ChartNet
ChartNet: A Million-Scale Multimodal Dataset for Chart Understanding 🌐 Homepage 📖 arXiv 📝 Changelog June 3, 2026 —…
ChartQA
Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…
ScienceQA
Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…
libero
This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase version": "v3.0", "robot type": "panda", "total episodes":…
commoncatalog-cc-by
Dataset Card for CommonCatalog CC-BY This dataset is a large collection of high-resolution Creative Common images (composed of…
stanford_cars
Stanford Cars Dataset Dataset Overview Splits: Training: 8144 images used for model training. Test: 8041 images used for…
OCRBench
Github Paper OCRBench has been accepted by Science China Information Sciences.
commoncatalog-cc-by-sa
Dataset Card for CommonCatalog CC-BY-SA This dataset is a large collection of high-resolution Creative Common images (composed of…
Innovator-VL-Instruct-46M
Innovator-VL-Instruct-46M Paper Code 🤗🤗 The data is being uploaded continuously Introduction To further enhance the model’s ability to…
GeoDrive-Bench
GeoDrive-Bench A multi-country driving scene benchmark for evaluating vision-language models on culture- and region-specific traffic knowledge.…