Dad, Dev, CEO, 4x Founder. Building @HelloUntangle

Connecticut, USA
So proud of my son for being a volunteer firefighter
14
129
6,331
My X creator payments go straight onto my @XMoney card and then Gill and I spend them on fun experiences. Enjoyed a nice breakfast this morning :)
15
69
8,116
Jev is very exciting
15
5
148
10,336
New @XMoney card arrived in the mail.
15
86
13,862
.@pulley is shutting down.
9
18
15,970
If you’re asking your team for progress updates, you’re doing it wrong. Your agent should have access to where the work is happening.
35
12
172
13,952
Running on @OmarchyLinux and asking agents to tune/improve your UX is just soooo good. Can’t imagine using macOS now. Wonder if @johnternus is paying attention.
53
4
370
27,468
I stop checking X for a day and I come back to everyone talking about fruit flies.
16
57
8,211
I agree with the sentiment here but I think David left out 3) @SpaceXAI - especially now that they include the @cursor_ai team 4) @Meta - it seems like Muse is poised to start really competing
Dario has written that we need to “pace the frontier,” and Sam has agreed. People may be surprised by my response: go ahead. You guys are the frontier. By any reasonable metric — market share, revenue growth, model capability — the two of you have a duopoly on frontier intelligence. You’ve also claimed the lead is widening because of recursive self-improvement. I don’t see what you see in the lab. If the unreleased models are scary enough that you think you should slow down, I support your decision to be responsible. But stop pretending you need anyone else’s permission. Stop pretending antitrust law has to be suspended so you can form a cartel. Stop pretending you need a regulatory approval process that supersedes product liability. Stop pretending METR is independent when it is intertwined with Anthropic’s investors and staff. Stop pretending you need those same evaluators to police competitors who aren’t even at the frontier. Most of all, stop pretending the motivation to slow down is purely altruistic. You face massive product-liability exposure if your products enable a truly damaging cyberattack. The market already punishes models that behave in unpredictable or unauthorized ways. After the Hugging Face episode, it is simply good business for OpenAI and Anthropic to trade some raw power for reliability and predictability. Call it alignment if you want. It is also just giving customers what they want. Pacing the frontier would also create breathing room for a more intelligent conversation about regulation than Bernie Sanders’ “shut it all down.” China is very unlikely to join a global agreement, as you know, and that has to be taken into account as well. So go ahead and pace the frontier. You are the ones setting it. The easiest way not to build superintelligence is for you to agree not to build it. Demanding your preferred regulatory framework as the price of that will look like blackmail of the public and the political system. So just do it. If you do, you’ll buy goodwill for the next conversation. If you don’t, we’ll know this was just another bid for regulatory capture — or an election-season psyop.
17
2
66
19,215
Ooooooh man
Go for launch
3
25
11,771
The “Embedded Evaluator” idea here is interesting
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
11
20
7,784
Nothing like having Gwyneth Paltrow read me Dario's latest essay :) Love @SpeechifyAI
6
17
6,481
We independently benchmarked Devin Fusion for its release today - this is the first time a multi-model coding agent has been included on the Artificial Analysis Coding Agent Index, and it effectively retains Claude Fable 5.1 and GPT-6 Astra performance while reducing costs Devin Fusion runs a frontier lead model with a cost-efficient sidekick. We tested configurations from Cognition combining frontier models from Anthropic and OpenAI with their new SWE-2 (medium) as a sidekick model. Configured with Claude Fable 5.1 (xhigh) + SWE-2 (medium), Devin Fusion scores 62 on the Coding Agent Index v1.5, while with GPT-6 Astra (xhigh) + SWE-2 (medium) it scores 59. The Fable configuration has the higher score, while the Astra configuration is 43% less expensive and completes tasks 31% faster. Congratulations to @cognition on the release! See below for our results and analysis 🧵
95
89
1,508
2,167,634
Been thinking a lot about 9/11 and all the brave souls who were killed and sacrificed their lives to save others 🇺🇸
4
55
6,378
The Cognition team gave me 50 free @DevinAI Max plans to give away. Each one is worth $200. Reply with the most useful or ambitious idea or project you’d have Devin ship for you this week. We’ll pick 50 people and DM the codes!
1,089
44
897
94,049
I would donate money to a non profit to secure computer and access to unreleased frontier models for the sole purpose of curing cancer. Does this already exist?
10
3
63
8,615
iPhone Duo is exactly the opposite of what I want: Less screens.
88
9
316
26,593
This post has 100,000,000 views
I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.
50
1
99
35,893