Meet Qwen-RobotSuite: Three Embodied AI Models for VLA Manipulation, Video World Modeling, and Navigation
The Qwen workforce has launched three embodied AI fashions, grouped as Qwen-Robot-Suite. The three are Qwen-RobotManip, Qwen-RobotWorld, and Qwen-RobotNav. Each is constructed on a Qwen vision-language spine and targets a distinct robotics downside. Qwen-RobotManip is a Vision-Language-Action mannequin for manipulation, constructed on Qwen3.5-4B. Qwen-RobotWorld is a language-conditioned video world mannequin with a 60-layer MMDiT and…
