π0.6 (pi-zero-6) is a vision-language-action (VLA) robotic foundation model developed by Physical Intelligence. It is the next version in the π line — a refinement of the π0.5 model on a slightly larger backbone — and serves as the decision layer (brain) of robots, including Ultra Robotics' stationary OP1 humanoid.
The π*0.6 (pi-star-0.6) variant is trained with the Recap method ("RL with Experience & Corrections via Advantage-conditioned Policies") — a three-stage process combining learning from demonstrations, expert corrections and autonomous reinforcement learning from the robot's own experience. Training on autonomous experience more than doubles throughput on the hardest tasks and cuts failure rates by at least 2x.
The model was demonstrated on complex, multi-hour real-world tasks: making coffee/espresso (running from 5:30am to 11:30pm), folding laundry (50 novel items in a new environment) and assembling 59 boxes on a real chocolate-packaging line, achieving over 90% success rates. The work was published on 17 November 2025.