Skip to content
Advertisement
Video

X-WAM-RoboTwin

X-WAM Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising Dataset Summary This is the RoboTwin 2.0 fine-tuning dataset…

X-WAM Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising Dataset Summary This is the RoboTwin 2.0 fine-tuning dataset used to train the X-WAM unified 4D World Action Model. It packages dual-arm bimanual manipulation demonstrations into a unified multi-view RGB-D video + low-dimensional state/action format, where each episode provides synchronized RGB videos, depth videos, dual-arm end-effector proprioception, actions, and a… See the full description on the dataset page:

Source: Hugging Face Hub (sharinka0715/X-WAM-RoboTwin). Metadata imported from the dataset’s Hub tags.

Advertisement