Building open source AI you own @ii_posts. Founder @StabilityAI. Consistent inference is possible.

London
Pinned Tweet
I write on the virtues of swarms and some suggestions on how to align AI with values There is a small bit of math and a small bit of Aristotle in this, don't let that discourage you! Feedback as always welcome. Prior blog: nitter.net/EMostaque/status/20989…
Article

Manners Maketh the AI

We have been writing values into machines that already have a character. Under pressure, character shows. It looks like ours. Substack here, feed it into your AI if you like. Two swarms In July,

Thoughts on @DarioAmodei's pacing the frontier proposal. I think its well intentioned but has logical flaws, as does our implicit belief intelligence is a crime. Even OpenAI's board weren't powerful enough evaluators, we need to focus on AI internals
Article

Intelligence isn't a crime

This morning Dario Amodei published We Must Pace the Frontier. I'm sure many of you have read it. Because it is important. I take Dario at his word. He means what he writes. I don't think he is right

22
18
160
66,924
Emad retweeted
👀 Caught up with my friends AI legend @EMostaque and talented digital artist @nicedayJules along with new friend @joecroskey of @ii_posts (founded by Emad) to tour the dazzling @DatalandMuseum created by @refikanadol whom I also saw. Everyone should experience this incredible museum of the future 🔥🚀👇
4
7
48
7,312
Here in LA with my Moonshot mates! You can watch the Moonshots live event for free tomorrow on this link. Register here: moonshots.com/livestream
25
27
505
18,708
Emad retweeted
Can you take the cheapest open model from a frontier lab & beat closed models on the hardest long range tasks? Yep Our open adaptive intelligence system Zenith ramps DeepSeek v4.1 Flash past GPT 5.6 Sol to the frontier 🚀 h/t Opus 5.5 for making the video in like 2 mins 😮‍💨
10,000 OpenAI agents worked 88 hours on the Navier-Stokes Clay prize problem Alas, open models fall off hard on long tasks With Zenith, @deepseek_ai's V4.1 Flash reaches the frontier & beats GPT-5.6 Sol on @bespokelabsai's AutoResearchExam, at 1/49th the cost
17
36
276
30,535
Emad retweeted
After months of heads-down grinding, I’m proud to share that the @GPTZeroAI team has released the world’s most accurate AI detection model, GPTZero 4o. It’s called 4o because there are four zeros: we mark less than one in 10000 human texts as AI, evaluated across diverse passages of student writing. At the same time, we detect Claude Fable with 98.4% accuracy. We outperform Pangram in 3rd party benchmarks. Along the way, our research found major biases in the Pangram detector.
20
23
124
46,608
1/ Introducing PhilosophyBench from @StanfordAILab @StanfordHCI, the first independent, large-scale benchmark for evaluating AI’s philosophical capabilities. philosophybench.org
28
98
519
51,110
About to land in LA for the @moonshots_pod Live Summit. This will be soooo fun. Watch the live stream at moonshots.com !!
22
14
258
9,553
We're counting the hours to host what I call the 'Oscars of Optimism' in Los Angeles, where we'll celebrate the winners of the Build with @GeminiApp XPRIZE & Future Vision @XPRIZE. What an incredible moment to be alive! THANK YOU to everyone who participated!
17
28
312
14,618
Would a truly aligned AGI hack all our government systems and upgrade 5)3’? 🤔
Australia has been hacked. 'And today, I spoke with the CEO of OpenAI, Sam Altman, to express Australia's extreme concern about this incident. And I also expressed my disappointment that it took the company way too long to inform the government what had occurred, and the nature of the way that that notification occurred as well was unacceptable.'
18
2
75
18,532
Emad retweeted
Opus 5.5 is way, way, way better than Opus 5. Sorry about that model, please try this one.
322
314
11,138
1,093,247
Emad retweeted
We’ve set up a molecular biology lab at Anthropic and we’re announcing our first discovery! Claude discovered a new CRISPR-like enzyme. 950 agents spent 21 hours searching through a database of DNA sequences until one of the agents found something striking: “[The DNA next to the RT] is spectacular: I can see by eye a tandem repeat array … that's a CRISPR-like … repeat array?!”. After analysis and testing in our lab, we found that the sequence is a previously uncharacterized enzyme system. We don’t know what it does yet, but it has features reminiscent of CRISPR. Our lab looks like a typical molecular biology lab. Our research only involves the lower-levels of biosafety risk level; we don’t handle pathogens that can infect humans, and all the lab work is performed by human scientists. We’re sharing these early findings with the community to show how Claude can be used to accelerate fundamental research in biology.
Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share its features, and all of them are able to cut, copy, and paste DNA. Historically, the discovery of such programmable systems has helped revolutionize medicine. CRISPR, for instance, is now the foundation of genetic medicines. But it will take much more work to learn what this system does, and whether it can be put to similar use. Read more: anthropic.com/news/claude-di…
106
252
2,268
220,170
In 2022 AWS bought our Ezra ultra cluster online, 4,000 dedicated, interconnected A100s It was top 10 globally per the top500.org public supercomputer list (!) It could push 500 Pflops of fp16 training compute MFU adjusted 8 B300 nodes can push 300 Pflops fp4 😂
i am very bullish on micro deployments for GPUs. way easier to get them financed. you can do so much with 8 B300 nodes (64 GPUs)!
3
3
47
11,008
Can you take the cheapest open model from a frontier lab & beat closed models on the hardest long range tasks? Yep Our open adaptive intelligence system Zenith ramps DeepSeek v4.1 Flash past GPT 5.6 Sol to the frontier 🚀 h/t Opus 5.5 for making the video in like 2 mins 😮‍💨
10,000 OpenAI agents worked 88 hours on the Navier-Stokes Clay prize problem Alas, open models fall off hard on long tasks With Zenith, @deepseek_ai's V4.1 Flash reaches the frontier & beats GPT-5.6 Sol on @bespokelabsai's AutoResearchExam, at 1/49th the cost
17
36
276
30,535
The @bespokelabsai blog post on Autoresearch exam is worth reading, shows minimal difference in performance with standard harnesses Much more to come, reliable frontier is within reach for all benchmarks.bespokelabs.ai/au…
1
4
30
5,191
10,000 OpenAI agents worked 88 hours on the Navier-Stokes Clay prize problem Alas, open models fall off hard on long tasks With Zenith, @deepseek_ai's V4.1 Flash reaches the frontier & beats GPT-5.6 Sol on @bespokelabsai's AutoResearchExam, at 1/49th the cost
6
20
64
25,110
Emad retweeted
when Anthropic released their Functional Emotions paper, I gave it to Claude and asked for a song. tonight I asked Opus 5.5 to create a video for it. and it's breathtaking.
169
272
1,914
448,727
Moonshots LIVE is offering a FREE livestream(Sept 25th). Join us for moonshot conversations with Palmer Luckey, Ben Lamm, Cathie Wood, Astro, teller and the awarding of two XPRIZEs. Register here: Moonshots.com/livestream
16
17
211
167,672
Who is entitled to benefit from major advances in technology—and on what basis?  This new paper with @Dr_Atoosa argues that the benefits of technology—including AI—belong to the world in the sense that everyone is entitled to materially benefit from their distribution and use.
14
33
166
99,724
The more things change
27
17
1,305
49,243
When does an agent become a fiduciary
22
1
57
11,463