Question Answering
190 results
Text
Measuring Massive Multitask Language Understanding
Measuring Massive Multitask Language Understanding
100K–1M·MIT·Parquet
ImageMultimodalText
megalith-mdqa
Images from Megalith, synthetically captioned using Moondream, with the questions then transformed to short-form QA using an LLM.
1M–10M·OpenRAIL·Parquet
ImageMultimodalText
MMStar
MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage 🤗 Dataset 🤗 Paper…
1K–10K·Parquet
ImageMultimodalText
ShareGPT4Video Captions Dataset Card
ShareGPT4Video Captions Dataset Card
10K–100K·CC-BY-NC·JSON
ImageMultimodalText
document-haystack
Document Haystack Dataset This repository contains the dataset for the paper “Document Haystack: A Long Context Multimodal Image/Document…