Skip to content
Advertisement
Image

MobileWorld-Eval-Results

MobileWorld-Eval-Results

MobileWorld-Eval-Results Evaluation results of the UI-MOPD trained model (Qwen3-VL-8B-Thinking) on a mobile agent benchmark. Contains full execution trajectories including screenshots, marked action visualizations, and task outcomes across 117 diverse mobile tasks. Evaluation Summary Metric Value Model Qwen3-VL-8B-Thinking Total Tasks 117 Successful 12 Success Rate 10.3% Action Space mobile use Max Steps 50 Avg Steps 32.5 Total Steps ~3.8K… See the full description on the dataset page:

Source: Hugging Face Hub (UI-MOPD/MobileWorld-Eval-Results). Metadata imported from the dataset’s Hub tags.

Advertisement