Skip to content
Advertisement
ImageMultimodalText

DocVQA

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval šŸ  Homepage…

Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval šŸ  Homepage šŸ“š Documentation šŸ¤— Huggingface Datasets This Dataset This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @article{mathew2020docvqa, title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)}… See the full description on the dataset page:

Source: Hugging Face Hub (lmms-lab/DocVQA). Metadata imported from the dataset’s Hub tags.

Advertisement