We had a blast last week when we hosted the first Hot Takes game night (+ dinner) focused on autoscaling RL envs. Carefully chosen guests were instructed to bring the most controversial opinions they could to a discussion on post-training.
As per usual, only technical practitioners, no VCs.
We had conversations on how to improve diversity of synthetic envs, distribution collapse, whether it’s even possible to do entirely autoscaling and RSI.
We host these specific research-focused events regularly, with different focuses and activities. If you’re working on any of the following problems, we’d love to include you in the next one:
- Unsupervised environment design
- Efficient RL training for multi-turn tool use
- Self-evolving benchmarks
- Autonomous AI research
- Open-Endedness