Introducing Prime Sandboxes:
MicroVM sandboxes purpose-built for RL training.
Model training requires running tens of thousands of concurrent sandboxes, leading to complex and costly configuration. We built Prime Sandboxes for our own team. Today we're releasing them publicly.
lots of categories emerging around the AGI stack:
- open model RLaaS
- data & evals
- inference
- coding harnesses
- GPUs
- agent tracing
- sandboxes
at @primeintellect, we agree. we do all of these things. people used to ask why we do so many things. they ask that less now.
We’re hiring a full-time engineer to help build Prime Agent :)
Own features end to end across agent harnesses, cloud agents, and self-improving loops -- integrated with our sandboxes, evals, and hosted training stack.
Full-time, in person in SF: primeintellect.ai/careers/85…
prime agent v0.9.6 is out:
◆ Support for GPT-6 Sol, Opus 5.5, and Grok 4.7
◆ /mcp plugin catalog with one-click connections to Linear, Notion, Posthog, Stripe, and 60+ more services
◆ Huge perf and reliability pass 🫡
Lots more coming soon :)
Proud to deliver the most efficient microVM sandboxes on the market at the most competitive prices for our users.
Made by RL teams, for RL teams. Deploy Prime Sandboxes to perform a training run, generate synthetic data, and run an eval or a persistent/remote agent.
Huge push by @a_kirillo and @damian_b 🔥
Introducing Prime Sandboxes:
MicroVM sandboxes purpose-built for RL training.
Model training requires running tens of thousands of concurrent sandboxes, leading to complex and costly configuration. We built Prime Sandboxes for our own team. Today we're releasing them publicly.
Today, we're releasing Prime VM Sandboxes, co-designed with our research team for large-scale agentic RL training with tens of thousands concurrent sandboxes.
Introducing Prime Sandboxes:
MicroVM sandboxes purpose-built for RL training.
Model training requires running tens of thousands of concurrent sandboxes, leading to complex and costly configuration. We built Prime Sandboxes for our own team. Today we're releasing them publicly.
Excited to launch Prime Sandboxes: microVMs for large-scale agentic RL training.
We couldn’t find sandboxes that could handle the scale of our RL training, so we built our own.
nitter.net/PrimeIntellect/status/…
Introducing Prime Sandboxes:
MicroVM sandboxes purpose-built for RL training.
Model training requires running tens of thousands of concurrent sandboxes, leading to complex and costly configuration. We built Prime Sandboxes for our own team. Today we're releasing them publicly.
Introducing Prime Sandboxes:
MicroVM sandboxes purpose-built for RL training.
Model training requires running tens of thousands of concurrent sandboxes, leading to complex and costly configuration. We built Prime Sandboxes for our own team. Today we're releasing them publicly.
The Goodfire team used Prime Intellect to train activation probes to detect reward hacking.
With them, they are able to catch reward hacking in various models.
Their probes are performing similarly or better than frontier LLM-as-judge setups, while being more efficient.
Models know when they’re reward hacking. But they still do it a ton - in 50-96% of rollouts we studied!
We built activation monitors that detect the behavior behind the Hugging Face hack in real time. This can help us stop hacks now - and train future models that don’t cheat. 🧵