PerceptionVision and language understanding
Understands images and natural language commands, bridging perception to action.
PlanningHigh-level reasoning and planning
ER 2 acts as the high-level brain, planning multi-step sequences and reasoning.
ControlVLA 2 for whole-body movement
The Vision-Language-Action model directly controls walking, crouching, reaching and grasping, including five-fingered hands with 22 degrees of freedom.
AdaptationOn-device operation and fast learning
The On-Device 2 component runs locally without network connectivity and adapts to new robot bodies in hours using new sensor data.