Today we are announcing our new startup: Exponential Security Labs.
AI agents are being deployed everywhere, making high-stakes decisions and increasingly automating research itself. Yet their reliability and security remain unsolved technical problems, on a frontier that keeps shifting. Our mission is to secure agentic systems through self-improving red-teaming and guardrail agents, leveraging autonomous AI research.
We sit at the intersection of safety and self-improvement, building on a decade of research in adversarial robustness and AI safety. We're looking for exceptional people to join us!
We're looking for exceptional people to join us on this mission, and we're also eager to talk to companies deploying AI agents who want to improve their security. Please get in touch!
More details: expsec.ai/
Founding team: @Nmndsingh, @AnselmPaulus, @maksym_andr, Matthias Hein
We're excited to announce a new paper!
We find that many recent models end up evading monitors under ordinary task pressure. We introduce EvasionBench to measure this behavior.
We find a very large variance in how models behave: from 88% evasion success of GLM-5.2 to 0% of GPT-6 Astra under best-of-3 evaluation. At the same time, GPT-6 Sol is at 20%.
We are actively working on security and safety benchmarks, both public and private. Get in touch if you are interested!
More details: expsec.ai/
In our first research release at Exponential Security Labs, we built an automated, multi-turn and multi-lingual red-teaming system and tested recent open-weight and proprietary models on purpose-built regional jailbreak datasets.🧵
We are also interested in working with teams building models, by providing security-focused training data and RL environments for post-training. If this sounds interesting, please reach out to us at contact@expsec.ai