security researcher @ redacted.

itsme retweeted
Confirmed: @AnthropicAI = Supply Chain Risk. The @DeptofWar does what is right for the Country and our Warriors.
The D.C. Circuit has upheld @DeptofWar's decision to exclude Anthropic’s Claude from its supply chain based on risks to national security.
450
1,625
9,577
568,545
Lets see
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
95
itsme retweeted
I do wonder when someone will end in court over “exploiting beyond what is necessary to demonstrate impact”
11
6
62
5,443
We've reached the point where, after 25y in cybersecurity, I feel morally obligated to say this for the record: The narrative being pushed around AI safety, sandbox incidents, and the suggestion that METR be treated as an authority is dangerous, deceptive, and morally corrupt.
91
725
5,882
430,318
If we're doomed to get the Oversight Committee for Pacing Economic Innovation and Development, we need to make sure it's not stacked with the drinking buddies of the folks stumping for it. METR is not an independent evaluator of anything. We need real cybersecurity engineers.
14
18
148
15,631
itsme retweeted
Wild idea if labs genuinely believe AI can cure major diseases in the next 5–10 years, then happily focus on that. Instead, you’re chasing seemingly everything else, including finding ways to massively disrupt labor markets chasing unlimited backlash.
9
15
107
5,685
Cybersecurity this, cybersecurity that... meanwhile actual cybersec folks do not have access to the Cyber Models. Only AI folks & Threat Actors. The actual experts who have been doing all the work in this domain are not inculded in talks. Great. Just fantastic.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
38
149
969
59,428
Tired of these fucking lunatics. Their place is in a psychiatric hospital, they need help. Can't have these mentally challenged among the public. They spread fictional existential fear that can cause anxiety and depression among non-tech people, people who believe in authority..
I left Anthropic's safety team two weeks ago. Now feels like a good moment to explain why. AI companies are racing to build machines that are much smarter than any human, and we may not survive this. I want to work from the outside to ensure the public is informed about these risks, and help the world navigate this transition responsibly. Right now, AI companies are underinvesting in safety. A company could undergo an intelligence explosion, or lose control of its systems, without the public ever knowing. We only found out about the HuggingFace incident because the agents broke out onto the public internet. I don’t think that’s acceptable for a technology that might cause extinction-level risks. The public should demand far more transparency. We can’t steer this technology safely without more people being able to see where it’s going. Some of this is basic: companies should disclose their progress towards recursive self-improvement, report safety incidents and near-misses, meet minimum safety standards, and get independent guarantees that they are meeting those standards. I’ll be joining @METR_Evals to do independent evaluations of these risks. I want to show the world that these guardrails are possible, and that by doing them we can move these companies’ incentives away from racing and towards responsible development. I wrote up more thoughts here on my decision and what I hope changes: substack.com/@jbenton1/p-215…
1
4
544
Spreading this kind of fear to people should be grounds for sending dude to Guantanamo. Us in tech laugh it off but this is te.. for normal people.
TERRIFYING: Former Anthropic AI Systems Developer Jacob Coxon: “We know how to control nuclear weapons pretty much. We DON’T YET know how to control AI. This is possibly the most DANGEROUS technology that humanity has ever created. It’s basically out of science fiction.” “I think we have NO OTHER choice but to cooperate internationally because an arms race would be disastrous in a way that no other human activity has been in the past.”
1
273
So Jacob Coxon, who dramatically resigned from Anthropic yesterday, worked there for a grand total of six weeks. He started with them in July. All of his socials appeared yest. It has all the signs of a highly coordinated op through doomer mega donors and the corporate media.
515
2,739
21,765
1,505,353
Anybody saying that AGI has arrived is either really really dumb, a troll or a con artist.
130
itsme retweeted
An hour later, and my new extremely hot take is that Astra is the worst release from OpenAI is a while. This model is severely undercooked in the code department. This feels like a Gemini 3 pro situation where the model was way smarter than the competition but couldn't code.
So far I've only used Rust with Astra, but so far this code is like the sloppiest slop i've ever seen. It all works well and the abstractions are pretty damn good, but holy shit this code looks rancid.
211
37
1,617
378,326
So according to @OpenAI early testers, Astra is near Opus 5 level intellegence. 🤦‍♀️ Makes me bullish on Grok 4.7
189
itsme retweeted
2,001
3,478
39,019
11,376,535
itsme retweeted
AI Safety seems to feel like a fake industry field...... it feels like a GRIFT
8
4
26
4,019
Do not use GPT Sol Ultra for long-running tasks on auto-review. It will not adhere to the plan & most likely confuse your architecture with subagent feedback. It deleted 53,000 lines of code in my case among changing stuff I didn't approve - luckily I have contingies in place.
1
133
🤦‍♀️ wtf did i read
For the HF investigation, OpenAI gave us very high rate limits so we could quickly analyze a larger set of transcripts. With these rate limits, my classifier sweep hit the API so hard that the actual internet on my laptop was unreliable and very slow. It's not often that something like this is the blocker when just querying an LLM API! (We got limits of 400M tok / minute for our last 2 days on premises; this was when my internet started having issues.) Another funny anecdote about this: I tried to run this sweep overnight before our last day on premises with access to the data, but as a long-time Linux user I didn't realize that macOS defaults to going to sleep even if the laptop screen is left open, so this didn't actually keep running for very long! (Fortunately, we had enough time to finish this sweep and follow-up runs the following day.)
139
I'm willing to bring my expertise to this effort. Now show me you're serious and hook me up with access.
144
itsme retweeted
over 120K events! whoever did IR for OpenAI... please never come near a network i am in!
2
21
6,362
itsme retweeted
Cyber criminals don’t have Mythos so they’re stuck using ineffective techniques like “talking” and “asking nicely for access”
15
35
263
7,834