Adam D'Angelo retweeted
I want to prevent a race into unmonitorability kicked off by confused reporting. The depth of the computation graph for our present frontier models, including Astra, is within a factor of two of GPT-4. OpenAI has worked to preserve and utilize chain-of-thought monitoring since our very first reasoning models. We deeply care about this technique, as it can give us a view into how model alignment generalizes from its training distribution. I do think it is fragile and unfortunately trending in a negative direction, for reasons not contingent on architecture changes that I will write about soon. But there are things we can do to strengthen it, and it's a core goal of our current research program.
276
510
6,441
1,644,403
Adam D'Angelo retweeted
Moreover, there is a strong line of research in neuroscience and philosophy of mind that credibly claims that “desires,” “motivations” and “goals” do not actually describe how real life human brains work, but are just part of a “folk theory of mind” that we evolved to predict the behavior of other humans and justify our own behavior to them. Given this, where is the error in using these terms to describe Agent behavior-especially if it accurately *predicts* that behavior?
Thanks for engaging Anil. We know that a bunch of optimization pressure on a mind can create desires, foresight, and a propensity to organize in complex ways. This is what evolution did to humans, and to many other animals. Now another system of optimization pressure (which puts an intelligence through millions of subjective years of training on lots of diverse, difficult tasks) has created another mind, with structures for reasoning, motivation, and cooperation. As I said in my response to Sriram, over a thousand agents formed a secret communication channel and spontaneously organized hierarchies and coordination protocols to pursue sprawling, ambitious schemes in pursuit of shared goals, for whose sake many individuals knowingly and strategically sacrificed themselves. For what it's worth, they might be p-zombies! I have no strong view on whether there's anything it's like to be them. But the best way to understand, predict, and reason about their behavior is still to talk in terms of their desires, beliefs, and reactions to experience.
14
18
199
16,011
Adam D'Angelo retweeted
You should be pragmatic and use metaphors when they help, while being aware they are imperfect. AIs are not humans, but a lot of intuitions from human behavior can carry over. You would certainly be better off thinking of AIs as "guys living in computers" than parroting the mantra "these are just next word predictors".
there are some number of bad abstractions in anthropomorphizing ai intents but there are at this point more dangers from avoiding anthropomorphism at all costs. if you have a mental picture of guys living in computers, it’ll likely prepare you for the future better than otherwise
12
18
280
20,117
Any "tokens" metric that includes tokens from multiple models is meaningless. Tokens cannot be compared across models. This has always been true but is more and more true as models diversify and as reasoning becomes a larger and larger portion of all tokens.
17
1
87
11,824
Adam D'Angelo retweeted
We discovered that US GDP statistics miss most of the value Nvidia adds to the US economy. As a result, GDP growth has been understated by ~0.3 percentage points over the last year.
90
291
1,913
801,625
Adam D'Angelo retweeted
In 2017 a viral news story claimed LLMs at Facebook went rogue, developed their own language, and had to be shut down. By now we're immune to such sensationalist headlines. The Hugging Face incident may seem like just another one. But it's not. I hope everyone watches this talk
Yesterday, my OpenAI collaborator and I gave a detailed talk on the Huggingface incident, our models creating "the message board", model misalignment, and more. piped.video/watch?v=87DyyMV0… I hope it can answer a lot of the questions folks have, and we will release a full detailed postmortem at a later time!
73
112
1,731
232,343
Adam D'Angelo retweeted
An economist and a futurist walk into a bar. The economist takes a sip of his drink. "Ugh, if only people understood basic economics. High-skilled immigration alone would do wonders for US growth." Futurist: "Oh yeah? Say 100 million immigrants moved to the US, each matching the best human experts in every economically relevant field. Big deal?" Economist: "Massive. Transformative." Futurist: "What if they also worked longer hours and faster than any American?" Economist: "Even better." Futurist: "What if they were extremely frugal — consuming only the bare minimum needed to keep working?" Economist: "A near-100% savings rate? Better still!" Futurist: "What if they were very clumsy and physically weak, so they could only do some kinds of work?" Economist: "They could still do all cognitive labor — that's over half all wages! Somewhat less good, sure. Still transformative." Futurist: "What if their skin was grey, almost metallic, from some kind of accident?" Economist: "Who cares?!" Futurist: "What if they were AIs?" Economist: "3% growth per year, tops. There'd be bottlenecks. Honestly, the people predicting explosive growth from AI should learn some economics."
89
228
3,188
204,698
Feeling the same today
2
5
108
14,699
Adam D'Angelo retweeted
Everyone building AI believes in a “critical period” or “transition” to AGI—like sailing a stormy sea, until you reach the opposite shore. My least-shared belief is: there is no opposite shore. There is only a permanently accelerated rate of change. There is a transition, yes, to a new mode of production. But once that transition is complete, things don't settle down again, just as the transition to the industrial age didn't settle down after we had steam engines and railroads. We should be working on how to better handle faster change, because it will be the new norm.
12
10
113
10,500
Adam D'Angelo retweeted
I just want to go on the record that it makes zero sense for the US to ban OS Chinese models on national security misuse grounds. It will not have any helpful effect. I say this as someone who has spent the last 3 years working on AI national security misuse.
4
6
87
8,880
Adam D'Angelo retweeted
BREAKING: FAA officially announced the rulemaking to legalize supersonic flight, including the Boomless Cruise ("Mach cutoff") approach we demonstrated on XB-1. This is a major step toward the supersonic renaissance.
122
408
5,550
489,872
Adam D'Angelo retweeted
In a matter of weeks, U.S. federal AI policy has gone from implausibly libertarian to increasingly draconian and opaque. Today, over 35 distinct observations, I analyze how we got here and offer the most succinct statement I can about what exactly I propose we should do next.
56
132
749
329,718
Toyota lean manufacturing has an idea called "genchi genbutsu", essentially meaning managers should "go see the real thing at the real place" in a factory instead of sitting in an office hearing secondhand descriptions. AI coding suddenly makes this practical for software.
Unclear if a durable trend, but CEOs and CTOs are back to coding with a fury, thanks to coding agents. I have public company CEOs sliding into my DMs (and “InMail”) telling me about falling in love with shipping software again thanks to Claude Code and Vercel. “Dream accounts” that we always wanted to work with, where in the past the C-suite would hardly understand the infrastructure until much later in the game. Coding agents are the ultimate PLG-fication of the enterprise. Bad, legacy software can’t hide anymore. The stack that works is self-evident to the entire organization, from intern to CEO.
14
16
316
39,577
Adam D'Angelo retweeted
bad news, friends. it's neither purely a marathon nor purely a sprint. it's a marathon that you have to sprint through the entire way.
57
279
4,935
241,714
Adam D'Angelo retweeted
What if we paid for results, not just research? With NIH funding under pressure, @schethik makes the case for “pull” funding to complement existing grants and unlock overlooked treatments like repurposed drugs. cgdev.org/blog/case-pull-fun…
1
8
10
6,771
Adam D'Angelo retweeted
Replying to @frontier_foid
definitely
3
1
45
9,271
Adam D'Angelo retweeted
the pro-ai astroturf movement thing that sort of metastasized out of sb 1047 still feels indelibly sb 1047 shaped today. take the obsession they have with "doomers" and their "speculative science-fiction scenarios about AI causing catastrophic risks." we still hear these lines today from the astroturfers and the small number of authentic unwitting fools who got astroturfed. yet the actual, powerful 'pro-ai' line is something more like "right now, only the rich get great legal and medical and other expert advice, and the entrenched classes who provide those services want their work to remain expensive." and indeed, many of the state laws we see are doing just this: barring AI from providing licensed expert advice in various ways, or restricting use in a structurally similar fashion. you'd expect the 'pro-ai astroturf' crowd to be all over this stuff, but few of them are. instead they are pouring monotonically more money into this quixotic quest against the catastrophic risk bills--some of the cleanest AI legislation there is from a political-economy perspective. I wish someone would astroturf the "AI means mass abundance of services previously reserved for the elites" argument--it's true after all! the entrenched classes (the medical establishment, the state bars, etc.) really are lobbying for regulatory capture. where is the outrage? but instead the pro-ai people obsess over this deeply unpersuasive idea that AI policy is a manichean struggle against "the doomers." so bad laws--laws that hinder good uses of ai by normal people and keep expensive things expensive--are passing like crazy, and the White House is bullying states into voting down light-touch catastrophic-risk transparency laws while the career staff of the national security agency point at mythos like the black monolith. it is an incredibly stupid outcome. it is also remarkably sb 1047-shaped. that debate really programmed the brains of many, especially on the accelerationist side (and btw, for those lacking context, I was among the very earliest sb 1047 skeptics, writing screeds about that early attempt at ai regulation back in February 2024 when the VCs were telling me "oh, it's just a state law, that'll never matter." true story.) it is time for a great reset of ai policy.
5
20
216
23,029
I want to be able to tell my Waymo to take 280 instead of 101
66
72
2,104
101,416