PolicyTrim speeds up VLA robots up to 5.83× without retraining
A Sichuan University team introduced PolicyTrim — a two-stage RL framework that cuts VLA robot task time by up to 5.83× in simulation and 1.86× on a real robot, with no training from scratch.