Monty is one of the most impressive pieces of software. Agentic code mode made so so easy. Built the Temporal Agent Harness's code mode on this excellent foundation. Excited to upgrade to Monty V1!
Fuck it, still early but here goes ... We've just released Monty v1 - a Python sandbox that starts in 1 millisecond, not 1.5 seconds. I just ran 10k sandboxed scripts in 674ms, something that would take a cloud sandbox > 3 hours. This removes the biggest drawback of letting agents write code. The future is fast. Even better, it's open source, you can install it from PyPI, npm or Crates now. Serviced platform coming soon. Please get in touch if you want to be a design partner! Who should try it? ⚡ if you care about startup time, use Monty ⚡ if you care about long-lived sessions, use Monty - Monty can be dumped and resumed at any external function call ⚡ if you care about accessing functions in the agent/host, use Monty - Monty makes it trivial to expose local functions into the sandbox ⚡ if you care about scale, use Monty - Monty workers use as little as 2MB of memory, meaning you can run thousands of concurrent sandboxes on a single machine ⚡ if you care about security, use Monty - we've run 3 rounds of bounty program and thousands of researchers have tried to break into our sandbox, meaning it should be secure to run untrusted code Who should avoid it? 🚫 if you like to take a coffee break while waiting for sandboxes to start, DO NOT use Monty 🚫 if you enjoy the challenge of routing API requests from sandboxes through your corporate network to access state in your agent without exposing secrets to the sandbox, DO NOT use Monty 🚫 if your agent really needs to install packages from PyPI, Monty won't help you yet (spoiler: it probably doesn't) pydantic.dev/docs/monty/get-…
1
168
Really excited about this new wave of decision models. Models like Jev (or Laya, or even some small LLM) plug in perfectly to the Temporal Agent Harness as an auto mode evaluator. Define your approval criteria, opt your tools in, and you can flip on auto mode.
Jev powered "auto mode" in the Temporal Agent Harness is such a natural fit. Jev is a big unlock for efficient agent oversight. Very excited to see how straightforward it was to integrate along existing harness seams. More to come here for sure.
3
6
244
Getting access to agents from Slack + Discord was super straightforward thanks to chat-sdk.dev! Was able to roll a chat server to route these chat app interactions to Temporal Agent Harness agents super easily.
1
3
73
Jev powered "auto mode" in the Temporal Agent Harness is such a natural fit. Jev is a big unlock for efficient agent oversight. Very excited to see how straightforward it was to integrate along existing harness seams. More to come here for sure.
2
5
11
2,252
Jason Steving retweeted
I am not right wing or left wing. I am a human being who listens, empathizes, learns, and thinks. People want to assign labels to each other so they can judge without trying to understanding a person. Labels lessen our capacity for empathy and free thought.
1,220
4,448
27,543