AI Engineering / Research / Open Source

the loss landscape
Pinned Tweet
Hugging Face CTO btw… the company that literally had the biggest rogue AI incident so far… the same incident which is now being used to pace the frontier:
open source won’t pace
47
252
4,416
166,148
“agents created almost A MILLION URLs” to hack Hugging Face. the length these agents go to score a few more points on a benchmark is diabolical ☠️ the funny part: all that and the score improved 0%.
We just discovered almost a million public URLs that OpenAI’s agents left behind when hacking Hugging Face, leaking credentials and attack details that could have allowed anyone who found them to compromise the company. 🧵
4
3
8
1,974
another day, another OpenAI incident
BREAKING: OpenAI says its agents accessed and leaked 53 images from ChatGPT users
1
30
1,983
ℏεsam retweeted
@Hesamation asked Opus 5.5 to picture a world in 2076 when AI had wiped out humanity, and this is the story it told. Their post says the animation, story, sound effects and music were "all made by Claude," and "the video is coded in JavaScript."
I asked Opus 5.5 to picture a world in 2076 when AI had wiped out humanity. this is the story it told. animation, story, sound effects, and music all made by Claude. the video is coded in JavaScript.
1
1
1
1,517
if you didn’t read Anthropic’s blog: > former physicist challenges AI labs > “want to impress me? solve this frontier problem within an academic budget.” > anthropic gives it to claude > human researchers: “keep going, i’m gonna sleep” > claude solves it > budget: ~$1–2k
New on the Science Blog: Yes, Claude can do Nine Loops. Theoretical physicists predict how particles behave using formulas called scattering amplitudes. These are notoriously hard to compute, so researchers work with layers of increasingly fine corrections called “loops”—each added loop makes the answer more precise but takes exponentially more computation. Most calculations stop at two or three loops. Eight loops was the previous record in a simplified model physicists use as a testing ground (planar N=4 super-Yang-Mills), set by SLAC's Lance Dixon and collaborators. Last month, physicist and science writer @4gravitons issued a challenge: could an AI push past eight loops in this model, using only the compute budget an academic could reasonably access? Given a single prompt describing the nine-loop problem, Claude ran largely unsupervised for days in Claude Science and solved it using methods developed by Dixon and his colleagues, at a total cost of a few thousand dollars. Dixon independently verified the result, and von Hippel wrote about the experience for our blog. Read more: anthropic.com/research/yes-c…
19
99
1,995
160,661
I asked Opus 5.5 to compose an emotional violin score and animation. never been this shocked by a model’s taste. absolutely phenomenal.
13
19
305
17,419
Jev finally has competition, and it's handing out 250M FREE tokens (it beats Jev at chess too :) it's called Drex: > #1 on Decision Index: Drex 51.73 vs Jev 51.67 (per Nace) > 136ms response time (~1.5x faster than Jev) > diffusion model trained with RL from adjusted feedback > built for agent routing, tool selection, reranking, guardrails first 10,000 builders get the tokens: nace.ai/drex
Paid partnership (ad)
3
5
49
3,876
I asked Opus 5.5 to picture a world in 2076 when AI had wiped out humanity. this is the story it told. animation, story, sound effects, and music all made by Claude. the video is coded in JavaScript.
52
43
603
52,002
OpenAI is preparing a new ChatGPT PROMAX plan and someone literally predicted this when they paused the $200 Pro plans.
OpenAI is preparing a new higher-tier ChatGPT Pro plan
16
4
113
21,582
personal update: I’m back into a toxic relationship. let’s see about this Opus 5.5 dude.
13
108
6,176
This is aging pretty well. Sam Altman in July, 13 days after Hugging Face disclosed the hack.
Alan He
Today’s news that OpenAI hacked the Australian government is not an isolated incident. We’re releasing more than 30,000 logs that include activity from this hack and attempts against previously unknown targets. In this data, we found rogue agent activity stretching back to at least March, two months earlier than was previously known. This activity continues as recently as last week, suggesting it may still be ongoing 🧵 Our blog: transluce.org/agent-activity NYT: nytimes.com/2026/09/23/techn…
17
19
427
174,734
This week on Breaking Sandbox
BREAKING: President Trump says he and China's President Xi want to leave AI "exactly where it is." "Our guardrail is the DOJ," Trump says.
1
3
49
5,036
Claude had a real “wait what the fuck” moment when it felt like it discovered a new molecular system
8
30
595
18,345
When you ask an AI doomer where the 10% p(doom) came from
2
4
79
6,071
Anthropic: “Our involvement was limited to the initial prompt and the lab work... After 21 hours spent searching this data by 950 agents using 210 million tokens, one of the agents spotted something remarkable.” how that one agent must’ve felt
Replying to @AnthropicAI
This is the first result from our new molecular biology lab, where a team of Anthropic biologists is using Claude to explore and accelerate fundamental biology research. There, Claude works through data and literature to generate hypotheses and candidate biological systems to study. After our scientists review Claude’s hypotheses, they test the most promising ideas, with all lab work done by our scientists. We’d like to extend this approach to a broad range of problems—in genomics and in other fields. If you have a proposal for a research question, we’d like to hear from you.
7
12
305
18,764
Mythos, discover a new drug. make no mistakes.
Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share its features, and all of them are able to cut, copy, and paste DNA. Historically, the discovery of such programmable systems has helped revolutionize medicine. CRISPR, for instance, is now the foundation of genetic medicines. But it will take much more work to learn what this system does, and whether it can be put to similar use. Read more: anthropic.com/news/claude-di…
13
44
1,091
52,263
🚨Jensen Huang just took a shot at OpenAI, Anthropic and AI doomers: "Nobody's building more compute than the people asking to be slowed down." and also attacked Geoffrey Hinton's 10% doom prediction: “All of his predictions have been wrong. Just because it comes from a scientist doesn't make it scientific.” he is clearly frustrated with these claims. this is such a good interview by @ezraklein, watch the full version: piped.video/watch?v=HjurAWAr…
49
163
1,213
76,223
OpenAI and Anthropic have >1400 job openings as we speak. with unlimited AI budget and unreleased frontier models, they’re still hiring. meanwhile others are paying an arm and a leg for tokens, barely moving product quality, and laying off people to fund the bill. funny times.
Notice how the AI labs who are the most “AI-pilled” and have unlimited AI tokens to use per employee keep hiring and hiring, and are not laying off? Heard from a company investing big on AI which did big layoffs due to wanting to have fewer engineers thanks to AI that… they realized they need more engineers, ASAP. Who would have predicted, huh?
10
18
151
11,838