Skip to content
Advertisement
Image

GQA-35k

GQA-35k

Dataset Card for GQA-35k The GQA (Visual Reasoning in the Real World) dataset is a large-scale visual question answering dataset that includes scene graph annotations for each image. This is a FiftyOne dataset with 35000 samples. Note: This is a 35,000 sample subset which does not contain questions, only the scene graph annotations as detection-level attributes. You can find the recipe notebook for creating the dataset here Installation If you haven’t already, install… See the full description on the dataset page:

Source: Hugging Face Hub (Voxel51/GQA-Scene-Graph). Metadata imported from the dataset’s Hub tags.

Advertisement