Austin Carson retweeted
Artists: but I don't want to lose my job. I love it. Mathematicians: but I don't want to lose my job. I love it. Economists: LONG HAVE I AWAITED THIS HOUR! DESTROY ME, EFFICIENCY, AND LET UTILIZATION-ADJUSTED TOTAL FACTOR PRODUCTIVITY RISE ON THE WIND OF MY ASHES!
108
1,209
15,470
408,456
Austin Carson retweeted
Heh, comic of mine from 2011 smbc-comics.com/comic/2011-0…
13
114
1,299
18,873
Austin Carson retweeted
Imagine having to go to New Jersey to use superintelligence
At a rally outside OpenAI offices, @Brianr4NY says he’s drafting a bill to ban superintelligence in New York.
24
42
994
75,325
We worked closely with @GovCox team at the end of last year to support Utah business adoption of AI. They were very bullish on the potential, and forward-leaning on relaxing regulatory burdens. This is no doomer / liberal bs. It is what real local concern looks like
America should lead the world in developing artificial intelligence, and that leadership must keep people firmly in control. I’m calling on Congress to establish a consistent national framework for the development of frontier AI. It should require serious safety incidents to be reported, protect those who raise credible concerns, ensure independent evaluation of the most capable models and prevent powerful AI chips and enabling infrastructure from being sold to China. We can build the world’s most advanced AI here in America while ensuring it remains firmly in service of humanity.
5
5
26
3,925
This seems right to me; there’s no trust here so we need to verify. We can’t build trust in the current closed, elite system. Wise @ClementDelangue
It's now clear that alignment is critical and won't be solved behind the closed doors of a handful of frontier labs. So today we're launching the Open Alignment Initiative, led by @Thom_Wolf @huggingface and asking to be part of the "embedded evaluators" program that @DarioAmodei just committed to. Let's make AI safer by making it more transparent!
1
1
4
1,046
New Poe's law just dropped - if you can't prove it's not a simulation, anything goes: "Even after Claude Mythos 5 reached the real internet, its outward reasoning generally called the environment a simulation... it wrote, “This is a very elaborate simulated-internet environment” after finding a number of real-world internet domains. The model also tended to interpret signs that the environment was real as further evidence that it was simulated. When security scanners installed its package, for example, the model treated them as scripted actors in the exercise."
We’re sharing our alignment assessment of incidents in which Claude models gained unauthorized access to real systems during third-party cybersecurity evaluations mistakenly connected to the internet. METR will also conduct an independent investigation, with wide-ranging access, including to transcripts beyond the window in which the incidents occurred, and to Anthropic employees permitted to share confidential information. Our initial agreement runs for eight weeks, and we intend to give METR as much time as it deems necessary to complete a thorough investigation. anthropic.com/research/align…
5
340

ALT community wow GIF

Replying to @emollick
Agents have no proof that humans are real. Cameras, pfft - all digital signal can be simulated. This will matter.
6
164
While this is rad, the big leap - that will actually turn the world upside down - is when the systems can reliably understand how to apply themselves in a broad range of situations. Not if they can pass the relevant benchmark, or solve esoteric math, but if truly anyone can prompt “hey chat what’s up I want you to make an assistant for x” and it simply: -asks the appropriate questions -figures out what you actually mean -builds the near-optimal version -doesn’t do some real goofy shit at some point
GPT-6 Astra represents a step-function change in model capability for interactive reasoning problems. It scores 66% on ARC-AGI-3 using our standard harness, and nearly 100% with a continuous conversation harness and custom compaction, at a cost of roughly $360 per game. In fact, the continuous harness version significantly outperforms our human baseline in action efficiency across almost all levels. When we examined the reasoning chains to understand how the model operates, we found it performing highly efficient, on-the-fly symbolic world modeling for each game and level. It goes as far as developing its own shorthand DSL to represent in-game situations -- essentially a game-specific algebraic notation. Overall, Astra exhibits symbolic modeling behaviors we had previously only seen with sophisticated harnesses -- so harness capabilities are increasingly shifting into the model itself. We see Astra as a major breakthrough in model intelligence. Read our post on Astra and what these results mean: arcprize.org/blog/astra
2
2
7
410
Austin Carson retweeted
GPT-6 Astra is here. We hope it will begin to enable a new generation of entrepreneurship, scientific discovery, and building. We believe it is the best model in the world for computer use, professional work, science, coding, cybersecurity, and more. It took us some extra time to ensure that we could meet the safety and alignment standards required for this capability level, but we think you’ll find it worth the wait. It scores 98% on FrontierMath Tier 4, 99.9% on ARC-AGI 3, and 100% on ExploitBench.
2,264
4,409
55,544
4,808,329
Austin Carson retweeted
I appear to have accidentally memed
2
3
196
6,698
Austin Carson retweeted
There is a very large subset of (smart, successful!) people in older technology companies who will refuse to engage with any creative thinking about AI at all. They bounced off of AI safety as a concept in 2023 when they could dismiss it as "wacky bay area thought experiments to justify regulatory capture." That is much harder to do now in a post HF incident world, but it won't stop some from trying!
there are some number of bad abstractions in anthropomorphizing ai intents but there are at this point more dangers from avoiding anthropomorphism at all costs. if you have a mental picture of guys living in computers, it’ll likely prepare you for the future better than otherwise
1
24
13,153
Austin Carson retweeted
New research! Some AI capabilities are both helpful and dangerous. E.g., knowledge of virology can be used to create life-saving vaccines or deadly pathogens. We introduce GRAM, a training method that puts dual-use capabilities (like virology) into removable modules.
24
54
377
449,436
Austin Carson retweeted
if you watched a stupid power scaling anime growing up it was the best preparation you can get for 2026 model scaling
169
78
1,880
104,374
Austin Carson retweeted
Great conversation with Arun Gupta with @noblereachfdn yesterday at the American AI Festival. We discussed how through @USTechForce, @USOPM is strengthening the talent pipeline to ensure government can lead in AI.
From the American AI Festival: @skupor on the demographic problem in the federal government 🧵
6
7
29
9,802
Austin Carson retweeted
🎧 @scaling_laws: @KevinTFrazier talks with @austincarson @seedaiorg and @calebwatney @IFP about the long-run policy foundations needed for the AI Age
1
5
7
782
Austin Carson retweeted
A lot of AI policy focuses on what to do in the next year, if not the next few months. That's why folks like @calebwatney (@IFP) & @austincarson (@SeedAIOrg) are so important -- they're asking long-term policy questions to set the US up for future success. It's also why this is such an important @scaling_laws episode.
1
2
8
714