Skip to content
Advertisement
Text

HR-Bench

Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models 🌐Homepage 📖 Paper 📊…

Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models 🌐Homepage 📖 Paper 📊 HR-Bench We find that the highest resolution in existing multimodal benchmarks is only 2K. To address the current lack of high-resolution multimodal benchmarks, we construct HR-Bench. HR-Bench consists two sub-tasks: Fine-grained Single-instance Perception (FSP) and Fine-grained Cross-instance Perception (FCP).… See the full description on the dataset page:

Source: Hugging Face Hub (DreamMr/HR-Bench). Metadata imported from the dataset’s Hub tags.

Advertisement