Security, agents and the systems behind them. Building, teaching and bringing builders together. @DevinAI Ambassador. Hybrid. | @pushcltv

๐Ÿ’ป
We live in bubble on this app. A lot of us are running agents, building workflows, creating skills, comparing harnesses, connecting tools, and etc.. Go ask one of your friends what Kimi or GLM is. They will probably look at you like a madman. OpenAIโ€™s consumer research found that 49% of ChatGPT messages are still people asking for advice or information. Another 40% are requests to complete a task. Meanwhile, 70.2% of sampled Codex users had assigned at least one task estimated to take a person more than an hour. Iโ€™m sure your timeline makes it feel like everyone is running agents and orchestrating workflows. They are not. Do not confuse being surrounded by early adopters with being late. You are early. Keep learning and keep building.
171
244
3,105
102,139
You asked for it, you got it. Devin TUI is officially open source๐Ÿ˜ฎโ€๐Ÿ’จ Runs on top of the Devin CLI you already have. Model picker with per model reasoning levels and pricing, Fusion, handoff to cloud Devin, session resume, @ file and image drops. Go break it and tell me whatโ€™s missing๐Ÿซก github.com/fenner888/devin-tโ€ฆ
Devin family, should I open source this or what?
6
3
35
1,478
You're going to pay the $500 GPT sub?!?

ALT Meme GIF by ALL SEEING EYES

10
1
22
839
Incase you missed.. Docker published a spec that turns an agent's permissions into something you can diff. A Sandbox Kit is an ordinary OCI image. The tools and the access they ask for ship together, so pinning the digest pins both. Their GitHub CLI example says it all. It asks for GitHub and nothing else. It can open pull requests but can't DELETE anything under /repos, because deny wins. And the token never enters the sandbox. The runtime injects it into requests to api.github.com, and the agent only ever sees a placeholder. A cool part: if a new version asks for another host or drops a deny rule, the runtime holds the upgrade and asks first. Widening authority stops being a routine update. Docker is clear about the limit too. A Kit grants itself nothing. Without a runtime that enforces the spec, the declarations are just an annotation. Permissions you can review before the agent ever runs.
Today we're introducing Cloud Sandboxes: same microVM isolation as Docker Sandboxes on your laptop, same CLI, always-on Docker-managed compute. Close your laptop. Keep the agent running. Move between local and cloud with one command. $250 credit for new accounts: docker.com/sbx-promo
3
12
684
Your agentโ€™s successful runs can become training data for a specialized model. LangChain launched LangSmith Fine Tuning yesterday, now in public beta. Its smithtune CLI takes curated agent sessions through dataset preparation, training with Fireworks or Baseten, evaluation and deployment. Those sessions include the messages, tool calls and results that led to an outcome. Youโ€™re teaching the model how to work through a task. In LangChainโ€™s internal code review evaluation, fine tuning Qwen 3.8 27B improved F1 from 48.9% to 53.7% while using roughly 30% fewer model calls. What caught my attention: an earlier, less selective training dataset actually made performance worse. They improved the data selection and tried again. I think you should judge this on held out tasks and cost per successful run. Training on your own agent history sounds promising, provided youโ€™re selective about which behavior you teach it.
5
1
13
420
The only 3 things I need to get this day started! Happy Friday all, ship something cool yet useful today โ˜•๏ธ
11
32
659
Devin family, should I open source this or what?
Word on the street is the CLI community was curious about Devin getting a TUI, so I built my own. @Chris_Wozniczek brought it up today on a call and got me curious about what it could look like. Obviously, I got straight to work. I think this would be clean. What do you guys think?
11
4
64
6,694
Me and @DevinAI after shipping 3 projects today.
7
32
695
Hybrid ๐Ÿƒ๐Ÿพ retweeted
Bro casually cooked a TUI for Devin. @markfenner Hey, are you okay? ๐Ÿ˜‚ insane!
Word on the street is the CLI community was curious about Devin getting a TUI, so I built my own. @Chris_Wozniczek brought it up today on a call and got me curious about what it could look like. Obviously, I got straight to work. I think this would be clean. What do you guys think?
1
2
3
202
Word on the street is the CLI community was curious about Devin getting a TUI, so I built my own. @Chris_Wozniczek brought it up today on a call and got me curious about what it could look like. Obviously, I got straight to work. I think this would be clean. What do you guys think?
13
1
45
6,704
One agent querying across AWS accounts without pulling every teamโ€™s data into one place? AWS has a reference design for that. Each team exposes a small set of MCP tools from its own account. AgentCore Gateway gives the agent one endpoint, and Cedar policies check the userโ€™s claims before a tool call. Think a lending policy lookup without copying the policy library into the central agent account. Tool results still travel back to the central agent for model inference. Keeping source data put does not keep every answer inside its original account.
3
12
432
Hugging Face is bringing llama.cppโ€™s ggml kernels into Transformers. I like this direction for local AI. This walkthrough shows how a GGUF model loads, picks up compatible Metal kernels, and generates locally through the familiar Transformers API. The fast path currently targets Apple Silicon and Qwen3.5-compatible models. Youโ€™ll need Transformers from main and compatible kernels.
1
7
297
Before letting Hermes edit your project, turn on checkpoints. Theyโ€™re off by default. Start your local CLI session with hermes chat --checkpoints. If a change goes sideways, use /rollback to list snapshots, /rollback diff 1 to inspect the changes, and /rollback 1 to restore that checkpoint. Enable it before you need it. Turning it on afterward wonโ€™t recover an earlier version.
5
17
499
Happy Thursday โ˜•๏ธ I think weโ€™ve all participated a few times in the Great Model switch of 2026. Now Iโ€™m conflicted ๐Ÿ˜ญ
11
22
567
Me and my main agent waiting for the subagents to finish its task.
18
1
42
1,230
Big shout out to the Devin community for the goodies. Building, sharing what I learn, and connecting with people along the way has been really rewarding. Thank you @cognition for the thoughtful gift. Appreciate you all ๐Ÿซถ๐Ÿพ
10
2
40
1,812
Okay, I tested Space Bunny Alpha. The experience was 10x better. One extremely short prompt. 6 minutes, 5 seconds. It built this playable neon arcade shooter. The cabinet design, colors and little details came together nicely. Just a first test, It's giving MiniMax vibes if you ask me..
Another stealth model. I hope this one is better than the experience I had with Union Alpha. Space Bunny Alpha just landed on OpenRouter. Itโ€™s a flash model with fast inference, adjustable reasoning, a 1M-token context window, and support for text, images, and video. Letโ€™s see!!
4
14
1,959