Skip to content
Advertisement
ImageMultimodalVideo

MMS-VPR: A Fine-Grained Multimodal Street-Level Visual Place Recognition Dataset and Evaluation Benc

MMS-VPR: A Fine-Grained Multimodal Street-Level Visual Place Recognition Dataset and Evaluation Benchmark for Dense Pedestrian Environments

MMS-VPR: A Fine-Grained Multimodal Street-Level Visual Place Recognition Dataset and Evaluation Benchmark for Dense Pedestrian Environments Overview MMS-VPR is the first large-scale multimodal street-level visual place recognition dataset featuring comprehensive integration of images, videos, and rich textual annotations with day–night coverage and a 7-year temporal span in dense pedestrian-only environments. MMS-VPR comprises 110,529 images and 2,527 video clips… See the full description on the dataset page:

Source: Hugging Face Hub (Yiwei-Ou/MMS-VPR). Metadata imported from the dataset’s Hub tags.

Advertisement