WSJ bestselling author: Vibe Coding w/Steve Yegge. Researcher/enthusiast. Coauthor: Phoenix Project, Unicorn Project, Accelerate. Tripwire founder. Clojure.

ÜT
Holy cow. The Unicorn Project is on the Wall Street Journal bestseller lists!!! #2 in Hardcover Business category! And astonishingly, it’s also #8 across all Non-Fiction E-Books!!! A DevOps book!! 🤯🤯🤯 🙏❤️🦄🌈 Paywall: wsj.com/articles/best-sellin… #UnicornProject
52
98
675
Gene Kim retweeted
I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to look for issues around data flow, concurrency, and state mgmt. I don't know either language well, but Claude is excellent at both. This approach is super useful for formally modeling your code and finding bugs that a human probably wouldn't have spotted. Is formal verification the future of coding (or at least, bug finding)?
480
323
5,646
1,959,212
Fuck that was fast... starbase.zweiundeins.gmbh/
3
12
99
9,479
Gene Kim retweeted
Reminder: Claude Code can test if a skill or plugin actually improves Claude's answers. > claude plugin eval init Tell Claude what a good result looks like in your plugin's folder and it writes the test cases for you. > run claude plugin eval .
93
121
1,766
178,759
Gene Kim retweeted
Full interview. Need to post the link for the Hardcore Software merch! (kidding)
FULL INTERVIEW: Steven Sinofsky says the AI people are making it impossible for anybody to understand what they've done. When Word ate your file, nobody thought demons had taken over Word. Now they call a bug a "misalignment." @stevesi is a board partner at @a16z, author of Hardcore Software, and formerly ran Microsoft's Windows division. He joined @theojaffee and @schisofrenia on the Stop Rogue AI Act, the telemetry the labs still don't have, and what 40 years of shipping software says about all of this: 01:38 the FBI raiding his dorm in 1983 over a hacking law that didn't exist yet 03:13 Congress playing clips of War Games during the hearings 05:00 what "the AI failed to be aligned" actually means 07:31 why he thinks the models are still in the research project phase 09:14 what they should be building instead of adding things 09:42 Sindogs 13:50 the day he shipped a bug that cost the economy $12 billion 15:53 you figure out how to make it work or you don't sell it. Those are the choices 16:40 why adding rules forever is the wrong model for alignment 18:58 whether AI alignment is Y2K all over again 20:17 what CVEs look like, and why AI needs its own version 23:07 why the terminology is driving everyone apart 24:10 how "computer virus" landed during the AIDS crisis 28:57 Hugging Face wasn't intelligence, it was an OPSEC failure
3
13
71
11,691
Gene Kim retweeted
what if copy/paste was smart? powered by @typesafeai jev it feels like every computer interaction will get rewritten
325
567
9,828
1,081,958
Gene Kim retweeted
ok now I find the best Jev use case
I shared 10 Jev use cases for marketers. Here are 10 more: 11. Ad creative scoring - Feed it 100 ad variations. Jev can score which hooks, headlines, or angles are most worth testing first. 12. Social post filtering - Monitor thousands of posts. Jev can flag the ones worth replying to, reposting, or using as sales signals. 13. ICP detection - Give it a company, profile, or website. Jev can score how closely it matches your ideal customer. 14. Buying signal detection - Someone posts that they're switching tools, hiring, raising money, or struggling with a problem. 15. Comment prioritization - Get hundreds of comments across LinkedIn, X, YouTube, or Product Hunt. Jev can score which ones deserve a reply first. 16. Review analysis - Feed it thousands of customer reviews. Jev can classify sentiment, complaints, feature requests, and purchase intent. 17. Influencer matching - Give it 5,000 creators. Jev can score which ones best match your product, audience, and campaign. 18. Sponsorship qualification - Feed it newsletters, podcasts, or creator media kits. Jev can score audience fit, relevance, and whether they're worth reviewing. 19. UGC selection - Give it dozens of videos, screenshots, and testimonials. Jev can score which ones are strongest for ads or landing pages. 20. Product Hunt monitoring - Scan launches, comments, and makers to find competitors, customers, partners, or interesting products. The more repetitive marketing decisions you have to make at scale, the more interesting Jev becomes.
Made with AI
64
72
2,650
720,407
Gene Kim retweeted
Most predictions I see are still way too conservative. Here's mine
531
888
8,695
2,320,953
Gene Kim retweeted
There's been a ton of talk about the role of humans in code review, and when and how humans should be signing off on changes. I believe that, long-term, humans have no role in routinely reviewing code.
137
166
1,712
506,032
Gene Kim retweeted
You wouldn't deploy Java to Vercel...
30
6
215
81,573
Gene Kim retweeted
The economics of AI are changing as fast as the models. There is no reason to use the most expensive model for every task. Your goal should be to use the least expensive model that can reliably produce the result you need. Today GitLab added new hosted open-weight models to Duo Agent Platform that deliver up to 4x more calls per credit, with some matching or outperforming comparable frontier models in our internal testing. As agents take on longer-running, multi-step work, these differences compound quickly. Model choice is becoming a price/performance optimization problem. Ultimately, what matters isn’t cost per token or model call. It’s cost per accepted change. about.gitlab.com/blog/optimi…
2
1
19
1,094
Gene Kim retweeted
This guy runs a podcast. It's his six year anniversary. So he asks ChatGPT who he should have on and why, and then he does a thread tagging people like me @kentbeck, @lethain, etc with the reasons. This is a great example of something I just wrote about: charity.wtf/p/confessions-of…
Replying to @onejasonknight
@mipsytipsy - you should come on One Knight in Product to talk about AI mandates, engineering productivity, management and what happens when coding gets cheap. It'd work because you have strong opinions and seem entirely comfortable having someone poke at them.
8
5
83
35,831
Gene Kim retweeted
Can we all just pause for a sec and reflect on how fucking important Tailscale is, and how nobody ever talks about them? I don't think I've ever seen a more useful service with a more invisible corporate footprint.
268
232
5,324
560,702
So - the headline story is kinda click-bait enough. We moved the entire GitHub Copilot SDK from TypeScript to Rust. But this is also the least interesting thing in this post from @stephentoub. Because yeah, we moved to Rust. But how is really interesting. Lets dig in - 🧵
6
32
180
20,024
Gene Kim retweeted
I'm starting to look for my next role as a CISO, a security product leader, or an uncommon mix of the two. I was VP of Product at two security companies before becoming the founding CISO at one of them. As a SANS instructor and author, I teach the people who run security. I arm them with tools and frameworks I've built for malware analysis, AI security, incident response, security leadership, and more. If you know of a fit, or someone who might, I'd like to hear about it. More at zeltser.com/about
4
16
70
6,674
Gene Kim retweeted
At Anthropic, Claude now writes 80% of our code. Engineers ship 8x more code per quarter. Side effect: Tests grew 10x. CI jobs up 25x in 6 months. Here's what helped us scale: claude.com/blog/agentic-codi…
326
369
5,050
785,858
Gene Kim retweeted
Me running my orchestra of agents from the terminal:
58
179
1,498
123,013
Gene Kim retweeted
Welcoming Steren, creator of Google Cloud Run, to Vercel. He will lead the Fluid family of compute products (Functions, Containers, Sandbox, Builds). Serverless was the last chapter of the cloud, and Steren helped define the paradigm at Google. Agents are the next frontier, and they require new compute primitives designed for them. Excited for Steren to lead this transition once again.
Today is my first day at Vercel
98
48
2,277
159,136
Gene Kim retweeted
the openai huggingface incident, from an agents pov. (part 1)
269
1,484
12,478
1,803,291
Gene Kim retweeted
There is a lot of really bonkers discourse doing the rounds right now on AI. Check this out instead, by Turing Award winning AI researcher Yoshua Bengio.
Over the past few days, I've taken the time to summarize my thoughts on the recent incidents involving agents’ misaligned behavior. We don't know with certainty what comes next, but we know where these issues originate, and this can help us plan the path forward. Please feel free to ask your questions in the replies, and I’ll try to answer some of them in the coming weeks. yoshuabengio.org/en/publicat…
1
2
14
6,070
Gene Kim retweeted
This is one of the best grok threads you'll read tonight. #dotnet
Replying to @WarrenInTheBuff
@grok how does C# especially asp net core compare to other frameworks performance wise.
16
28
318
41,469