Skip to content
Advertisement
ImageMultimodalSensor / Time-seriesText

AutoGUI-v1-702k

This is the training set of AutoGUI paper AutoGUI: Scaling GUI Grounding with Automatic Functionality Annotations from LLMs ✨We are glad to see that…

This is the training set of AutoGUI paper AutoGUI: Scaling GUI Grounding with Automatic Functionality Annotations from LLMs ✨We are glad to see that our dataset is adopted by top-level GUI Agents, such as ByteDance-Seed/UI-TARS and Step-GUI. Data Fields Each sample in the dataset is either a functionality grounding or captioning task. “image” (PIL.Image): The UI screenshot of this task. Note that the images are at various resolutions. “func” (str): the functionality annotation of… See the full description on the dataset page:

Source: Hugging Face Hub (AutoGUI/AutoGUI-v1-702k). Metadata imported from the dataset’s Hub tags.

Advertisement