Open
2,922 results
360Motion-Dataset
360°-Motion Dataset Project page Paper Code Acknowledgments We thank Jinwen Cao, Yisong Guo, Haowen Ji, Jichao Wang, and…
fine-t2i
Fine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning arxiv by Xu Ma, Yitian Zhang, Qihua…
OmniDocBench
OmniDocBench English 简体中文 OmniDocBench is an evaluation dataset for diverse document parsing in real-world scenarios, with the following…
commoncatalog-cc-by-sa
Dataset Card for CommonCatalog CC-BY-SA This dataset is a large collection of high-resolution Creative Common images (composed of…
Innovator-VL-Instruct-46M
Innovator-VL-Instruct-46M Paper Code 🤗🤗 The data is being uploaded continuously Introduction To further enhance the model’s ability to…
document-haystack
Document Haystack Dataset This repository contains the dataset for the paper “Document Haystack: A Long Context Multimodal Image/Document…
kb-books
open-rdl-books Dataset Description Language dan, dansk, Danish License Public Domain, cc0-1.0 Dataset Summary Documents from the Royal Danish…
GeoDrive-Bench
GeoDrive-Bench A multi-country driving scene benchmark for evaluating vision-language models on culture- and region-specific traffic knowledge.…
MMMU
This is a merged version of MMMU/MMMU with all subsets concatenated. Large-scale Multi-modality Models Evaluation Suite Accelerating the…
DocVQA
Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage…