SynthUX Visual Computer-Use Trajectories
SynthUX Visual Computer-Use Trajectories
748 results
SynthUX Visual Computer-Use Trajectories
SciReC: Diagnostic Evaluation of Relational Reasoning in Multimodal Scientific Conversations with Adaptive Interaction
WorldMemArena WorldMemArena is a large-scale multimodal memory benchmark designed to evaluate how well AI systems retain, update, and…
Vietnamese Medicinal Herb VQA
SynDataLab/Irodori-Ja-Spk2-10k 10,000 single-speaker conversational Japanese utterances synthesized with Irodori-TTS-500M-v2. Part of a 4-speaker…
Irodori TTS Reference Voices v2 (10K) 10,000 reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign (no ref=True) using a richer…
Bengali TTS — Missing Rows This dataset contains the rows from rwd51/bengali-tts-combined that are not present in the…
Bengali Indic TTS Dataset This dataset is derived from the Indic TTS Database project, specifically using the Bengali…
Bambara-ASR-All Audio Dataset
A systematic and standardized benchmark for the multi-dialect Arabic zero-shot TTS task Paper: "Habibi: Laying the Open-Source Foundation…
InfoRe Technology public dataset №1
Lahgtna Levantine TTS — Synthetic Levantine Arabic & Code-Switching
Urdu-Munch-Lina Processed version of zuhri025/Urdu-Munch with LinaCodec encoding. Dataset Structure This dataset contains 2 batches of audio data…
Bashkort TTS Dataset The largest open dataset for speech synthesis in the Bashkir language — featuring multi-speaker recordings…
Kahwa Postcast Arabic Speech Dataset
wuwa-voice-EN wuwa voice EN is a dataset of voice line from Wuthering Waves Attribute Value Language English Total…
Syrian Postcast Arabic Speech Dataset