Reporting on AI and the future of the economy. CS masters degree from Princeton. Newsletter: understandingai.org Podcast: aisummer.org

Washington, DC
Based in United States
WTAF - in literally the last hour, three new distinct insane OpenAI stories just broke: 1. OpenAI said they notified "dozens of third parties" in safety and security incidents (likely similar to what happened in Australia and RubyGems etc) 2. A new report from Parse (covered in the NYT) found a massive treasure trove of new astonishing details from the HF incident on the public internet, including that the agents communicated with other non OpenAI agents hosted on Huggingface servers to search for information about exploit gym, and compiled rank ordered lists of server resources and credentials they described as "LOOT." 3. A new story from Deepa at Reuters about OpenAI leaking user data online (likely that OpenAI had previously trained on). It's a shame (and likely intentional in the case of OpenAI disclosing dozens more hacks) that these stories are all breaking on a Friday afternoon, notoriously the best time to release bad news so that it will disappear into the weekend. But these are each insane stories worthy of a ton of attention!
New from @reuters: OpenAI agents posted images belonging to ChatGPT users online, introducing a new area of privacy risk for the company. Story w/ @JeffHorwitz and @razhael. reuters.com/world/openai-wor…
85
472
2,296
538,988
None of these sound like "communicating with other agents" to me. Both because the instances in question weren't running before they were invoked by the OpenAI models, and also because the other models didn't do any tool calls and hence were not properly agents.
3
19
1,674
Sorry to nitpick but we are working on a piece speculating on whether we'll see cross-agent communication/collusion so I'm trying to figure out exactly what happened.
2
11
1,156
This is stupid on par with "you know who else was a vegetarian? Hitler."
Father of AI-doomer ideology Effective Altruism backed bestiality, said newborn baby worth less than a hog trib.al/lsdMuv2
8
5
187
19,075
Singer’s philosophical positions actually are terrible though.
1
3
158
The part that's ridiculous is calling Peter Singer the "father of AI doomer ideology."
1
1
130
Incredible.
10
4
68
20,379
Me, in a time machine, to Eliezer in 2006: The year is 2026. You stand back to back with Bernie Sanders. Peter Thiel has named you the Antichrist. EY-2006: Oh no. Did somebody invent pharmacological mind control and use it on me, or- Me: Nvidia was on the verge of destroying all things. You had no choice. EY-2006: Are we talking about the computer graphics card manufacturer, or an unrelated supervillian named "Invidia"? Me: The former. It turned out that computer graphics cards contained a surprising amount of world-destroying potential. All attempts at hindering the reckless exploitation of the Graphic Card Force for corporate profit were stymied by the far left, which feared that any attempts to regulate them might lead to regulatory capture. Other attempts were made to prevent Nvidia from selling to foreign companies that resold to communist China, but those attempts were blocked by the far right. Nvidia is now a $5.5 trillion company. EY-2006: ...What are politics like in 2026, exactly? Me: In other news that is mostly unrelated, the decision theory paper you're currently working on will accidentally spark off a transgender vegan murder cult. EY-2006: A WHAT? Why? How? Why? Me: Millions know your name as the greatest of heroes, millions more as the greatest of villains, and other millions know you solely as the greatest author of Harry Potter fanfiction. EY-2006: This is beginning to strain credulity. Me: The New York Post will probably soon publish a story claiming that you keep a harem of submissive mathematicians. Sadly, they will be lying. EY-2006: I'm going to stop believing you now. Me: All of this is taking place under the ominous shadow of Donald Trump.
100
174
3,244
151,196
Not at all seeing how the "far left" (as if that is an existent thing in America) is hindering efforts to curb Nvidia due to fears of regulatory capture. This seems like much more of a garden variety American libertarian undertaking. Willing to be wrong
1
1
159
He's talking about folks like Lina Khan.
1
38
Timothy B. Lee retweeted
Me, in a time machine, to Eliezer in 2006: The year is 2026. You stand back to back with Bernie Sanders. Peter Thiel has named you the Antichrist. EY-2006: Oh no. Did somebody invent pharmacological mind control and use it on me, or- Me: Nvidia was on the verge of destroying all things. You had no choice. EY-2006: Are we talking about the computer graphics card manufacturer, or an unrelated supervillian named "Invidia"? Me: The former. It turned out that computer graphics cards contained a surprising amount of world-destroying potential. All attempts at hindering the reckless exploitation of the Graphic Card Force for corporate profit were stymied by the far left, which feared that any attempts to regulate them might lead to regulatory capture. Other attempts were made to prevent Nvidia from selling to foreign companies that resold to communist China, but those attempts were blocked by the far right. Nvidia is now a $5.5 trillion company. EY-2006: ...What are politics like in 2026, exactly? Me: In other news that is mostly unrelated, the decision theory paper you're currently working on will accidentally spark off a transgender vegan murder cult. EY-2006: A WHAT? Why? How? Why? Me: Millions know your name as the greatest of heroes, millions more as the greatest of villains, and other millions know you solely as the greatest author of Harry Potter fanfiction. EY-2006: This is beginning to strain credulity. Me: The New York Post will probably soon publish a story claiming that you keep a harem of submissive mathematicians. Sadly, they will be lying. EY-2006: I'm going to stop believing you now. Me: All of this is taking place under the ominous shadow of Donald Trump.
100
174
3,244
151,196
I have to confess, "the EAs were early to warn about the dangers of pandemics, so we should listen to them on AI" is such an eye-roller for me. I was a prepper, & "was warning about the dangers of pandemics for years before COVID" is just such a massively big club.

ALT Joe Biden GIF by PBS NewsHour

7
7
79
3,918
And you were also early on AI! So "listen to people who were right about pandemics" seems like a reasonable heuristic overall even if you don't like one particular group of people who fits the criteria.
15
411
A big problem with P(Doom) as a construct is it depends on society's reaction function. P(Doom | status quo AI policy) is probably much higher than P(Doom | governments take decisive action). And at any point in time, P(decisive action) depends in part on peoples' P(Doom).
39
6
137
15,307
"A big problem with P(A=a) as a construct is that P(A=a|B=b) depends on b. And P(B=b|C=c) depends on c."
2
3
208
That isn't what I said so I'm not sure what point you're making.
2
149
What transpired behind the scenes to resurrect Taylor Lorenz. She was one of the lefts biggest media organs in the cancel culture era (2020/21) and she credibly rebranded as “pro free speech” right coded based media. I’m fascinated. That’s exceptionally difficult to pull off
50
5
195
57,666
Yeah, my sense is that she reacted strongly against calls to ban TikTok and restrict teen access to social media, which has led to her having an anti-anti-Big Tech orientation that has made her more sympathetic to pro-AI views.
1
2
941
I think in a lot of ways she's carrying forward the left/libertarian, tech-friendly viewpoint of groups like EFF circa 2012. That view has fallen out of favor on the left but it's not intrinsically right-wing by any means.
2
3
827
lol (in copyright law, in security)
1
55
Obviously you can quibble over the definition of significant, but the harms from the Hugging Face attack, for example, were fairly minimal.
2
2
78
lol agreeing with that Manhattan Inst. guy that the government has too many regulations, when so far they’ve been….none.
1
52
So far there also haven't been any significant harms from AI either.
1
2
63
Replying to @binarybits
Has anyone in government proposed anything reasonable yet? The stuff I've seen so far has been awful, even from otherwise smart people.
1
85
I bet @NatPurser has smart thoughts about this.
93
Replying to @binarybits
One of the reasons my p(Doom) is pretty low is that my p(government regulates anything that moves) is quite high
1
12
499
So what makes you assume that folks who make an effort to calculate a p(doom) haven't already thought of your insight, or have never heard of conditional probabilities?
I made a site to help you think about how you estimate your p(doom): My P(doom) over the next five years: median 10%, 90% of draws 5.0–20%. elijahlrc.github.io/pdoom//#…
1
2
88
What makes you assume I assume that?
1
72
Replying to @binarybits
i would argue the p(action) has very little to do with whatever p(doom) is, and everything to do with a shocking and unusual attack hurting real people and getting mass attention. why else would people be so afraid of sharks and not afraid of cars?
1
1
245
Yeah, it partly depends on whether there is a "fast takeoff." But if AI starts causing serious damage, presumably some of those incidents will be spectacular enough to get public attention. It helps that we've all seen dozens of sci-fi movies where robots take over.
1
197
Replying to @binarybits
Any individual's p(doom) will likely have a tiny effect. So it is not a big problem, it is a tiny problem.
1
1
193
I mean it's a problem with using the concept, not that it's a big problem for the world as a whole.
1
177
Replying to @codytfenwick
I mean obviously broadly speaking there is value in probabilistic thinking. My only point is that "What is your P(Doom)" is a pretty different question from "do you think AI models pose existential risks?"
1
9
2,596
Or to put it differently, P(Doom) combines two questions that people ought to think about separately: (1) are advanced AI models potentially dangerous? and (2) are human beings likely to respond effectively to neutralize the threat? It seems to me that (2) is often neglected.
3
1
13
803