Unitree open-sources UnifoLM-WLA humanoid foundation model
Unitree Robotics has released UnifoLM-WLA-1.0, an open-source 6-billion-parameter vision-language-action (VLA) foundation model designed for whole-body humanoid robot control. Built upon Qwen3-VL-4B and trained on approximately 2,500 hours of real-robot demonstrations, it coordinates 64 tabletop and whole-body manipulation tasks across diverse end effectors within a unified action space.
Open-sourcing whole-body humanoid VLA models bridges the gap between high-level vision-language reasoning and low-level physical dexterity, accelerating robotics research beyond proprietary ecosystems. Unified action spaces provide cross-embodiment versatility by allowing a single policy to switch between parallel grippers and five-finger dexterous hands without model architecture rewrites. Furthermore, dynamic region prediction enables interaction-centric world modeling so the policy anticipates scene changes before generating trajectories, while 2,500 hours of real-robot demonstration data drastically mitigates common sim-to-real transfer failures.
DISCOVERED
1h ago
2026-09-13
PUBLISHED
1h ago
2026-09-13
RELEVANCE
AUTHOR
AI Search