Skip to content
Advertisement
MultimodalTabularText

Hy-Embodied-0.5-VLA-Data

Hy-Embodied-0.5-VLA-Data

Hy-Embodied-0.5-VLA From Vision-Language-Action Models to a Real-World Robot Learning Stack Tencent Robotics X × Tencent Hy Team 📖 Abstract We introduce Hy-Embodied-0.5-VLA (Hy-VLA) — an end-to-end Vision-Language-Action system that spans the full robot learning stack: data collection, model design, pre-training, supervised fine-tuning, RL post-training, and real-world deployment. Built on the Hy-Embodied-0.5 MoT backbone, Hy-VLA integrates a flow-matching… See the full description on the dataset page:

Source: Hugging Face Hub (tencent/Hy-Embodied-0.5-VLA-Data). Metadata imported from the dataset’s Hub tags.

Advertisement