nuScenes
1,000 driving scenes with 360° camera, LiDAR and radar.
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks Abstract: Vision-Language Models (VLMs) have demonstrated effective…
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks Abstract: Vision-Language Models (VLMs) have demonstrated effective perception and reasoning capabilities on general-domain tasks, leading to growing interest in their application to Earth observation. However, a systematic benchmark for comprehensively evaluating remote sensing vision-language models (RSVLMs) remains lacking. To address this gap, we introduce OmniEarth, a benchmark for evaluating… See the full description on the dataset page:
Source: Hugging Face Hub (sjeeudd/OmniEarth). Metadata imported from the dataset’s Hub tags.