Creator and maintainer of zvm.app | software engineer | writer at shippingbinaries.com

United States
Pinned Tweet
785
The fat cat is so wordy but also very good at catching itself. No support for high reasoning on @vercel for some reason.
21
20% you say… 🤔 what other optimization mathematically can only result in a 20% increase.
Do we actually have better software since AI? What company, what product got 2x better in the past 3 years? Vibecoding feels faster but the production gain is 20% at most
27
TanStack is quaking
a weird inversion with LLMs is the models improve faster than the tinkerers when i see people with custom workflows and setups they're all addressing problems that don't exist anymore the person naively using vanilla codex is more likely to be experiencing state of the art
3
578
Hook up Claude with the Drive connector and you’ve got the God prompt.
Finally! 🥰
52
Why didn’t I think of that?! This is the benchmark for system one models.
Replying to @alokbishoyi97
Here's a fun benchmark we ran: 8 decision models playing Tetris. In our tests @perplexity_ai Decider came consistently at the top !! ( cc @AravSrinivas ) you can try em out here - app.routerplus.com/playgroun…
27
I really like this idea. Gotta design it, but I think I’ll support this in ZVM.
I've published a new, general-purpose terminal specification that gives any program a way to tell the terminal what it's doing: idle, working, waiting, finished, or failed, and why. It is easy to implement on both sides. I've written more about why: mitchellh.com/writing/progra… The problem this solves is generic and applies in compelling ways to all sorts of use cases (e.g. package managers, build tools), but it really rears a hideous head when it comes to AI and the "agentic inbox" problem. I found over 250 different agent orchestrators (with various goals) that all individually implement a heuristic based approach to detecting whether things like Claude Code are working, blocked, done, etc. If you look at @herdrdev's commit history, you can see around 10 compatibility changes to detect Claude Code within just the past 3 months. (This isn't a dig at Herdr! They're doing great work with what they have, and everyone else is doing the same thing). Heuristics are not the way. And over-indexing on agentic use cases isn't the way, because this problem is really interesting for other tools too: Homebrew, Terraform, Cargo, etc. etc. Proprietary protocols and out of band APIs are not the way. They have an O(N) integration problem and require extra work to work over SSH or within local VMs/containers. The program status specification (OSC 7501) solves this in a reliable, deterministic way that is general-purpose and can be easily adopted by the entire industry. Read more: mitchellh.com/writing/progra…
1
1
40
I’ve yet to see a project that has meaningfully improved via carcinisation “Oh my app is so fast. My app is so safe.” You are building an HTTP server to serve a CRUD app.
1
169
I’ll raise you open source medical software for cancer/health/science research.
there has never been a more altruistic and inspiring project for a SWE to work on than america.gov
2
2
144
Atalocke retweeted
just one more agent bro
2
44
545
9,983
And you use it to write TypeScript.
For decades, researchers have sought materials that sort electrons by spin while their magnetism cancels. In 3 days, 90+ Opus 5.5 agents helped us uncover two room-temperature magnetic semiconductor candidates in simulations: YBaMnFeO₅ and KV[Cr(CN)₆]. KV[Cr(CN)₆] was synthesized back in 1999. Its predicted ability to sort electrons by spin appears to have been hiding in plain sight for 27 years.
2
78
Atalocke retweeted
Just so we are all on the same page: * The Rust version has hundreds of commits, with multiple optimizations rounds, over several days. The others have 1 or 2 commits from an initial agentic rewrite * Based on the available documentation, different things were asked from each attempt. The agents.md from Go mentions the Rust version was the canonical example. The Elixir one mentions it was the Rails one * While we don’t have the full prompts, it seems they were vague, allowing the agents to make non-language architectural decisions that impact results (such as how closely it should replicate Rails’ drawbacks) It is actually amazing how quickly we can have full-blown rewrites up and running. But if we want to take conclusions from the data presented, it is important to understand the process behind it.
The agents finished porting Campfire to Laravel and Django too. I'm sure there are optimizations that could be found there too, but the whole point is to get a sense of how frontier agents do with different environments out the box. github.com/basecamp/once-cam…
39
86
1,109
51,171
Did this with my Fedora desktop. So nice to not have to dig into CUPS when I just need to print a doc.
AGI achieved internally. Claude just got my wireless printer to work with my desktop. Mathematicians since Babbage have debated whether this is possible. The question is now settled.
42
That’s called a book.
We created gyms because modern day work no longer required physical activity. Without exercise, our muscles atrophy. I predict we’ll need gyms for our brains too, once AI starts doing our knowledge work.
2
40
Atalocke retweeted
A framework is an intelligence cache. You could waste tokens reinventing the ORM, cache, queues, auth, routing, etc, or you could just pick a world-class framework.
I don’t know if I agree with this take. Good frameworks are intelligence caches. The production of verifiable high-quality code has a cost to it (insert for now caveat). Frameworks can lower this cost as long as they don’t get in the way of agents.
79
86
1,070
48,164
Muse, create a dashboard showing me all of the dashboards I have.
39
“Ack” is such an ugly word. It sounds like my cat about to throw up. 🤢
32
Atalocke retweeted
This process is mandatory and cant be bypassed
290
4,723
142,370
4,925,776
Atalocke retweeted
waking up and seeing your agent stopped to ask for permission 8h ago
311
2,544
45,422
1,015,996
Best 51 seconds of tech news I've seen today.
example.com just changed for the first time since 2013, and IANA replied to my email asking why. Here's the full response: oliverdunk.com/2026/09/30/ia…
1
1
83