steerable and explainable AI

San Francisco
We're hosting an after-party at Spark Social SF for @ycombinator Startup School 2026 builders! Come grab some drinks + snacks and meet the d_model team. 📍 Spark Social SF 🗓️  Saturday, July 25 | 6–8pm PT 🎟️ Register → luma.com/jz0h1t7q Registration is on a first come, first serve basis. Everyone who registers is entered into our raffle with the chance to win gift cards and d_model swag. We'll draw winners at the event, so show up for your chance to win!
1
701
d_model retweeted
Replying to @d_model_ai
We're working directly with frontier-lab researchers to do this work with bleeding-edge models. We expect to produce many more blog posts like this as our research program grows symbiotically with our RL-env work. And we're hiring :) dmodel.ai
2
6
355
Discovering Concept-Editing Algorithms With LLM Agents dmodel.ai/concept-erasure/
3
7
20
2,918
I've shifted my research to focus on automated alignment research. We will have automated AI research very soon and it's important that alignment can keep up during the intelligence explosion.
We estimate that, on our tasks, Claude Opus 4.5 has a 50%-time horizon of around 4 hrs 49 mins (95% confidence interval of 1 hr 49 mins to 20 hrs 25 mins). While we're still working through evaluations for other recent models, this is our highest published time horizon to date.
73
118
1,369
456,044
d_model retweeted
no its not
This robot solving a rubiks cube in 0.103 seconds is a little preview of what "AGI" really means
10
2
152
4,422
I'm doing research @d_model_ai now! Very excited about the team and mission. We're teaching computers to teach computers to teach computers to be good people.
1
1
6
410
d_model retweeted
Replying to @d_model_ai
Mfw parted illusions overshadow marinade
1
2
525
what are some good accounts to follow for someone just getting into tweeting about ai
1
4
786
maybe i can get gemini to run this account
Replying to @TheZvi
Good at fiction writing and surprisingly eager to do it, without the self-conscious Assistant breaking the fourth wall all the time. Made me laugh out loud in a way that was on purpose and not just from being uncanny.
2
624
goals are supposed to be lofty
Replying to @cremieuxrecueil
Let's drop the lofty goal that any of humanity's vices will be left behind just because we hop in a rocket to go somewhere else.
1
1
5
784
cant believe this guy is replying to someone saying lets try to not do slavery again
2
208
can someone give me an example of this pls
2
1
21
2,383
new hire just asked for more equity and less cash during negotiations, huge W
1
27
9,150
d_model retweeted
the true value in AI coding nowadays isn’t in writing code, it’s in reading code: outlining large code bases, navigating, retrieving suggestions & ideas, and breaking down and explaining tricky concepts i’ve never had a more complete understanding of codebases before
I am learning that I am an unusual CEO in many ways. Some know that I still code on nights and weekends (github.com/lattner) despite a busy “day job”. As part of that, I care deeply about sw development as a profession, and what “AI anxiety” is doing to many early and experienced developers. IMO, AI is here to stay, super useful today for things, and will continue to improve… but is far from SWE job displacement. It is sad for me to see many talented people giving up hope and sabotaging their own career development because of a future that may not actually arrive! Please watch the video for a much more nuanced discussion about the issues involved, particularly if you are an early career engineer who wants to make a big impact on things!!
4
3
52
7,346
d_model retweeted
📢 Can LLMs really reason outside the box in math? Or are they just remixing familiar strategies? Remember DeepSeek R1, o1 have impressed us on Olympiad-level math but also they were failing at simple arithmetic 😬 We built a benchmark to find out → OMEGA Ω 📐 💥 We found that although very powerful, RL struggles to compose skills and to innovate new strategies that were not seen during training. 👇 work w. @UCBerkeley @allen_ai A thread on what we learned 🧵
23
151
711
183,816
Do LLMs actually understand the code they write? 1/ We show that programming concepts like “nullability” can be directly extracted from the latent representations of language models. 🧵
3
12
85
23,649