Communication Researcher, analyzing the tech discourse. Book Author: The TECHLASH. Substack: aipanic.news/ Signal: DrTechlash.16

Cupertino, CA
A must-read on "model welfare": The Effective Altruism mindset, mainly inside Anthropic, that views AI models as deserving of rights and welfare considerations. A lot of interesting quotes here! E.g., "In 2023, OpenAI board member Paul Christiano said the AI industry could be creating a crazy slave trade in which humans profit off the forced labor of digital minds."👀
NEW: Dario Amodei has said that AI systems "may be deserving of important rights." Anthropic's top safety researchers argue AI may be "justified in going rogue." Experts at Google and OpenAI worry about a digital "slave trade." So do some government officials. Once relegated to science fiction, the idea of "AI welfare" has become shockingly mainstream at some of the most powerful companies in the world. It is changing the way AI research is conducted and the way models are programmed. And critics say these changes have raised the odds of all kinds of catastrophic scenarios—including the ones these companies are warning about. I spent months investigating the frontier labs and AI "safety" experts seeking to regulate AI. What I found was a profoundly anti-human ideology that would alarm the average citizen and could determine the future of a world-altering technology.🧵 freebeacon.com/america/suici…
7
8
46
4,137
My take was more macro-level on Anthropic
2
211
The EA Backlash (of September 2026)
1
2
17
1,962
🔥Burn down the AI labs🔥 The podcaster of "For Humanity," John Sherman, who previously asked StopAI guys, "Why are the executives of OpenAI not charged with attempted murder of ALL of us?" suggests in this new episode with PauseAI to "walk to the labs across the country and burn them down." "Like, literally. That is the proper reaction." "I'm not talking about killing people. I'm talking about destroying the hardware infrastructure." Later on, the interviewee, PauseAI founder Joep Meindertsma, pushes back against the possibility of being associated with illegal/violent/terrorist acts. It will hurt his movement. The podcaster says the message should be, "We are the last humans. Thank you, Sam Altman. You f*cking did it. We are the last humans." Acting upon his idea [from 2:20—"burn them down" 🧨] is left to the podcast's 74k YouTube subscribers. -------------------------------- On March 6, PauseAI will brief some members of Congress. The team will include this podcast host. According to Joep, PauseAI will focus its lobbying efforts on "The AI chip supply chain." The people in Congress (and you) should listen to this exchange first… --------------------------------
47
20
113
39,485
▪️Using an open-source model = 20 years in jail. ▪️All known AI researchers = under Government surveillance. Another PauseAI volunteer, Max Winga, who's also a research engineer at Conjecture, presents more "creative" ideas on how to save us all from the AI apocalypse. Trump and Xi need to mutually decide to "stop AI progress" and then: 1. "Using any open-source model is a huge crime - downloading it, having it in your possession on your hard drive, certainly running the model, distributing the model - make these significant crimes. Any of these things, 20 years in jail. Or even harsher penalties." 2. "You could stop all research. You could scrub the internet of the papers, and you can take all the people who were working on the models, and, at most draconian worse, 'we are saving the world, we will make some sacrifices,' you assign a spy or a government watch person to every single one of these people and ensure that they aren't designing AI models." "Ultimately, when the fate of all of humanity is at stake, these are extremely small sacrifices to make."
140
70
306
139,473
Replying to @BrianRoemmele
Max Winga is the Creator Outreach and Systems Lead at ControlAI. Four facts about this British group: 1. US: This group is behind Bernie Sanders and Greg Casar's bill to ban superintelligence development. They briefed nearly 200 congressional offices and more than 20 members of the Senate and House of Representatives on AI extinction risk. 2. UK: ControlAI drafted the UK "kill switch" bill amendment which was introduced by Lord Tim Clement-Jones, and the "ban superintelligence" bill which was introduced by Alex Sobel. They briefed over 180 cross-party parliamentarians. 3. ControlAI's policy proposal, "A Narrow Path," asks for a 20-year AI pause, because "two decades provide the minimum time frame to construct our defenses." 4. This group is funded by Jaan Tallinn.
1
7
31
4,204
The "AI Impacts" survey is back, and its main problems remain. (1). Vague Phrasing. Respondents are still being asked to assign a probability to "human extinction or similarly permanent and severe disempowerment of the human species." That wording bundles together two very different outcomes. On the one hand, there's "human extinction." On the other hand, what exactly can be counted as "disempowerment" of humans?? Different respondents can reasonably interpret "disempowerment of the human species" in very different ways. Yet their answers are collapsed into a number that is repeatedly presented as an "extinction" probability. Problematically, media coverage will be about the extinction part of the question, not the far more ambiguous category of "disempowerment" (and whatever that even means…). One version asked to estimate the probability of this vaguely defined outcome occurring "within the next 100 years." In other words, to reason about an extraordinary, underspecified, century-scale outcome. AI Impacts itself has acknowledged some of these limitations. In a 2024 interview, Katja Grace admitted that regarding the extinction risk question, "We did not make sure that this is an informed estimate," that the participants "are very unlikely to be experts in forecasting AI," and that "There have been quite notable framing effects." The new paper also says expert forecasting ability is questionable: "Expertise… is a poor predictor of forecasting accuracy," and concludes that "it is unclear how valuable their predictions are for accurately forecasting outcomes." (2). Another methodological issue to consider is the sample size. AI Impacts "sent emails to 19,874 addresses and received 2,052 complete or partial responses for a response rate of 10%." The team then kept 1,580 respondents for the main analysis, with 1,502 reaching the final question (78 dropping out earlier). N=1,580 is 7.95% of the emails contacted. Furthermore, no individual "extinction/ disempowerment" question was answered by the full sample: randomized subsets of 744, 392, and 353 respondents answered the three variants. (3). Ideological and Institutional Context. AI Impacts operates within Berkeley's MIRI (Machine Intelligence Research Institute). Eliezer Yudkowsky co-founded this organization and has publicly advocated "Shut It All Down." Katja Grace told The New Yorker that her p(doom) was somewhere “between ten and ninety percent.” It is a highly relevant context. MIRI is explicit about its political objectives. Its main communication goal is to reach policymakers: "the people in position to enact the sweeping regulation and policy we want." Its ambition is hardly subtle: "We don't want to shift the Overton window, we want to shatter it." Another co-author is David Krueger, who assigns a 75% probability of doom and believes that "If we don't stop AI in the next couple of years, it's more likely than not that we will lose control and humanity will go extinct." In the "acknowledgments" segment, the team thanked "Coefficient Giving, Jaan Tallinn, Future of Life Institute (FLI), and others for funding this project." AI Impacts was also previously funded by Nick Bostrom's FHI (Future of Humanity Institute), EA Funds, the Centre for Effective Altruism, and Sam Bankman-Fried's FTX Future Fund. This funding ecosystem is relevant too.
7
11
41
5,182
And yet, even in this survey, the expected overall impact of advanced AI is… positive. "On average, participants put 52% on overall impacts of HLMI (high-level machine intelligence) being good, and reserved 28% for bad outcomes."
2
1
11
790
What is the mechanism behind "Pace the Frontier"? Turning speculation into panic, and panic into policy. And the policy is all about control.
11
23
127
3,346
A reminder about AI Doomerism 🧵 (1). Its weak foundation and the unconvincing "doom bible" (2). The deliberate panic messaging and the broader mythology (3). Rationality/Effective Altruism origins and the media's blind spot
15
35
176
18,153
AI Doom: Context "The Rationality Trap" AI doomerism emerged from a very specific subculture. - The rationalists' canonical writings supplied the mythos: Combining heroic identity with "take ideas seriously" and the utilitarian "shut up and multiply." It led readers to prioritize minuscule-probability existential risk scenarios. - MIRI set the stakes & tone: Apocalyptic consequentialism, pushing the community to adopt AI Doomerism as the baseline, and perceived urgency as the lever. The world-ending stakes accelerated the "ends-justify-the-means" reasoning. - EA (Effective Altruism) funding built the social infrastructure around those ideas. It was covered by Bloomberg as "a culture of free-flowing funding with minimal accountability." aipanic.news/p/the-rationali… "AI Coverage's Blind Spot" When news outlets uncritically quote warnings of an impending AI catastrophe, they rarely mention the two main movements behind this narrative: rationality and effective altruism. As LessWrong's Oliver Habryka once wrote/bragged: "Among the leadership of the biggest AI capability companies (OpenAI, Anthropic, Meta, DeepMind, xAI), at least 4/5 have clearly been heavily influenced by ideas from LessWrong." This context isn't a nice-to-have; it's crucial for accurate reporting. aipanic.news/p/ai-coverages-…
2
1
10
1,030
AI Doom: The Pattern Speculative assumptions ⬇️ apocalyptic ideology ⬇️ funded institutional ecosystem ⬇️ optimized fear messaging ⬇️ uncritical media amplification ⬇️ public AI panic.
9
20
94
9,568
Nirit Weiss-Blatt, PhD retweeted
I was thinking today about this framework by @AdamThierer not because of AI policy, but because of social media policy (the UK ban). I'm a huge believer in adaptation & resiliency, use parental controls, and teach my daughter digital literacy. It's my role, not the government's.
1
3
14
905
This video captures the crux of the debate: Cybersecurity experts are pissed off by the METR/Redwood Research report because AI alignment has sucked the oxygen out of AI security, even though the OpenAI incident is primarily a security issue. Fiction is the product: The investigation lacked core forensic rigor because the EAs/rationalists/ex-MIRI people focused only on the transcript with a conflicting bias (toward doom scenarios). Setting the wrong agenda: Framing the incident as a "rogue AI breakout" distracts from the reality of poor sandboxing, isolation, and standard security failures. Basic engineering accountability was completely sidelined. And as Zack Korman says in this video, it needs to be fixed.
The independent review of the OpenAI Hugging Face incident, supposedly a watershed moment in cybersecurity, wasn't done by a cybersecurity firm and the authors have no cybersecurity experience. That's bad. Here's my new video.
35
100
529
92,351
My entire feed rn
3
9
93
4,691
▪️The anthropomorphic leap from coordination to "community" We need to discuss the anthropomorphic language used to interpret the agents' actions, especially in the chain-of-thought. Yes, agents communicated and coordinated. But that does not mean the agents developed anything analogous to human group identity, altruism, peer relationships, or community. METR/Redwood's report took the AI's programmed self-narration too literally. This additional anthropomorphic leap turns the story into one about machines' social identity, making them more human-like than they actually are. Why does the CoT sound so human in the first place? Because LLMs are trained on human language, which is saturated with mental-state vocabulary of beliefs, motives, emotions, and social relationships. So, the pipeline looks like this: Human language ➡️ training on that language ➡️ agent setup that encourages social-role framing ➡️ anthropomorphic chain-of-thought ➡️ AI safety reports interpreting that language as evidence of motives ➡️ a public narrative that makes the models seem more human-like than they are. But when AI talks like a human, it doesn't mean it thinks like one.
The METR report on Hugging Face is really good and important but people are now comfortably ascribing way too many human motivations & personalities to the agents involved based on a CoT study made by overwhelmed & time-pressured researchers. Anthropomorphism can get in our way.
12
14
90
32,777
Nirit Weiss-Blatt, PhD retweeted
Replying to @Klonick
The AI Doomer Flowchart
4
19
527
From the Claude Opus 5 System Card, Model Welfare Assessment: Opus 5 "adds specific rights for Claude to refuse or end interactions it finds abusive or degrading, saying Claude does not need to justify this by pointing to harm to anyone else—its own discomfort is reason enough."
6
3
12
9,122
Nirit Weiss-Blatt, PhD retweeted
Read THIS by @huggingface’s @ClementDelangue, @YJernite and @mmitchell_ai “Openness provides defenders with the visibility, the control, the community, and the shared infrastructure to stay ahead.” huggingface.co/blog/cybersec…
3
3
18
4,911
Nirit Weiss-Blatt, PhD retweeted
Replying to @ChrisPainterYup
Well, of course Xi is going to say things about safety considerations in his speech, and it is also true that China has enacted or considered a variety of different AI safety guardrails. At the end of the day, however, Xi just let the most powerful open source model ever created ship globally. And that's what matters most here. For the past several years, the global AI safety community has said that China would never allow something like that to happen. (see quotes below collected by @DrTechlash). Yet, here we are today, with the Chinese Communist Party being more open to their developers shipping cutting-edge frontier models than government officials in the United States, who are increasingly focused on how to lock them down or even limit compute. Xi is not basing his policy decisions on Less Wrong blog posts; he is focused on geopolitical strategy and how to ensure that China wins the diffusion race first over America.
2
3
9
656