releasing many high quality open-source RL environments is the most impactful thing anyone can do to push the open-source frontier right now
the equivalent of sharing high quality pretraining data but in the new RLVR paradigm
the most insane part, they will release ~7k RL training data and the framework leading to this top 6 model on AA, they also shipped the model + tech report less than 1 week after starting the final RL run
pushing both intelligence and openness level, huge congrats