Skip to content
Advertisement
ImageMultimodalText

BLINK

BLINK: Multimodal Large Language Models Can See but Not Perceive 🌐 Homepage 💻 Code 📖 Paper 📖 arXiv 🔗 Eval AI This page contains the benchmark dataset…

BLINK: Multimodal Large Language Models Can See but Not Perceive 🌐 Homepage 💻 Code 📖 Paper 📖 arXiv 🔗 Eval AI This page contains the benchmark dataset for the paper “BLINK: Multimodal Large Language Models Can See but Not Perceive” Introduction We introduce BLINK, a new benchmark for multimodal language models (LLMs) that focuses on core visual perception abilities not found in other evaluations. Most of the BLINK tasks can be solved by humans “within a… See the full description on the dataset page:

Source: Hugging Face Hub (BLINK-Benchmark/BLINK). Metadata imported from the dataset’s Hub tags.

Advertisement