this essentially proves that a massive human teleop data is not necessary for robotics. once we reach RSI(which we likely have), the model only requires a VLA sandbox to continuously run sim-RL, which enables it to resolve and grasp physical principles so robots can operate in the real world.
GPT-6 Astra scored 95% on a robot control task, up from Fable 5.1's 40%, with 6.2x fewer output tokens at 2.3x lower cost. 🧵