Skip to content
Advertisement
ImageMultimodalText

FineBench

Dataset Card for FineBench FineBench is a large-scale, multiple-choice Video Question Answering (VQA) dataset designed specifically to evaluate the…

Dataset Card for FineBench FineBench is a large-scale, multiple-choice Video Question Answering (VQA) dataset designed specifically to evaluate the fine-grained understanding of human actions in videos. It leverages the dense spatial (bounding boxes) and temporal (timestamps) annotations from the AVA v2.2 dataset, providing ~200k questions focused on nuanced person movements, interactions, and object manipulations within long video contexts. Dataset Details… See the full description on the dataset page:

Source: Hugging Face Hub (FINEBENCH/FineBench). Metadata imported from the dataset’s Hub tags.

Advertisement