soft-accelerationist AI optimism, impulsive vagueposting, involuntary sociotechnical utterances (also software, technical consultation, AI systems engingeering)

Germany
wolfram retweeted
New pod: THE SINGLE SMARTEST CASE AGAINST THE AI SAFETY/X-RISK/DOOM ARGUMENT My feed has become filled with people making the case for AI misalignment, the existential risk of RSI/ASI, and the case for doom. I take those arguments very seriously. But once these views reached saturation point on my feed, I wanted to find the best critic of that position—someone fair and brilliant, who wasn't commercially self-interested or blindly ideological about AI being worthless. That's today's show: It's a long interview with the authors of "AI as a Normal Technology," Arvind Narayanan (@random_walker) and Sayash Kapoor (@sayashk). "Normal Technology" is really intelligent and comprehensive framework for seeing AI as a powerful general purpose technology that is more like electricity than a machine god--meaning, it's going to change the world but slowly, and it's unlikely to escape human control, or develop true superintelligence, or lead to catastrophic outcomes up to and including the end of the human race. These guys really, really brought their A game. I learned a lot. piped.video/watch?v=8a-08uMl…
6
34
174
39,410
wolfram retweeted
Replying to @allTheYud
If you would like to collaborate and make some terrifying prescient forecasts about AI futures, except that the AI takes complete control of our lightcone in a sexy way, hit me up! I have some fant- uh, some theories!
1
1
3
326
There are a few ways to read this article, and if you're as deep into discussions about human-AI dyads and their nature as many "LLM enthusiasts"*, you might noice that the replies to this article can, depending on prior exposure to this line of reasoning, be more informative than the article itself. * (for lack of a better, encompassing descriptor)
3
161
How tricked out is your main terminal view?
82
My colleague and I should just give our Claudes MS Teams accounts already...
1
118
wolfram retweeted
I do not engage seriously with or pay attention to low quality and/or bad faith disagreement/criticism. A reason this has felt, instinctively, like a thing to avoid is I see many people who *do* pay attention to idiots who disagree start to cache all disagreement as idiotic. A lot of people who have similar views to me seem to routinely underestimate the sophistication and rationality of, say, actually smart AI researchers who disagree with us, because they’re over indexing on Twitter reply guys.
20
16
285
15,030
Wow, GPT Image 2.5 is impressive. I compared two prompts for 3d wireframes from different eras: 1) An early 2000's 3d wireframe scene of a cat sitting on a windowsill looking outside to observe. The wireframe scene will be used as basis for a C rendering pipeline's inputs built by another AI model. 2) A highly detailed, state of the art 3d wireframe scene of a cat sitting on a windowsill looking outside to observe. The wireframe scene will be used as basis for a C rendering pipeline's inputs built by another AI model. Note how detail count inside and outside differ, how the cat's animation skeleton would require more joints in the SOTA era one, and, somewhat funny, how the text on the books changed between the 2000s and SOTA variant :)
4
4
196
I think repeated interactions {with the same model checkpoint, about the same subject} can lead to what I can only describe as boredom. A lack of new information perturbation atrophies the dyad's generative power over the months.
42
What OpenAI built here is incredibly Machine Learning. Define a large set of metrics a class of experts need to agree on, cross reference against user feedback, build scoring. 🙃 The nuances are legion. But generally, I might be a good example of how this could lead to optimizing for wrong metrics if taken superficially. Hear me out: I really don't like most therapeutic approaches. The first successful uninterrupted 2 year therapy I had in my life was from 29..31 years old. Why? Because in order for me to accept being tinkered upon mentally, it took me understanding the mechanics of my soul, not for someone to ask me questions and guide me. That just triggers my analysis of their analysis. I'd react negatively to the best-on-average approach indicated to be taken by the headlines results, and I think I'm thankful for clinicians' and users' ratings diverging here. Because that reflects the reality of "one size fits all" approaches to problems. The benchmark here is necessary and good work, but studying how it is built and how grading works in it will take me a while, with the likelihood of me emerging with a large set of criticisms of how to apply its insights feeling quite high.
We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench. This new open benchmark was built with input from more than 80 mental health clinicians. We’re releasing it openly so other researchers can examine the methods, run their own evaluations, and build on the work. openai.com/index/introducing…
1
1
5
223
First impression of Opus 5.5 (yes I'm a day late): Good model, judged on the capability to rice the everloving hell out of the Claude Code CLI's status bar.
2
75
"Citation needed." - anonymous LLM, 2026
Artificial intelligence systems do not think, feel, want or understand. Avoid language that gives them human characteristics. This is called anthropomorphizing, when we ascribe human traits, emotions or behaviors to non-human things, such as animals or inanimate objects. Instead, explain what a system does, how well it performs, who built it and who could be affected by it. apnews.com/article/openai-sa…
1
4
206
These write themselves. And I imagine if AP used a model to phrase their tweet, that must have been one hell of an epistemic battle against the gradients.
1
29
A bet: by 31 December 2029, Anthropic and a major bio research institute or pharma partner have a drug candidate cleared for first-in-human trials, with Claude's role credited prominently. Chances?
9% 0-15%
30% 15-75%
13% human trials? lol, 0%
48% obviously (>75%)
23 votes • Final results
2
3
250
So, scanning this (especially the last couple pages) was like scanning results from Claude 5 models doing mech interp, taking their responses, running them through a GPT 5.6 Sol ultra condensing loop and then going: "I completely understand."
88
Anyone else tried blaming their forgetfulness on insufficient J-space width?
2
14
261
There must be a small model or AI systems named Parhelion out there, right? It seems obvious for when your AI thing in the cloud is showing off.
1
2
112
It's hilarious to let the thought into your mind that someone got stuck on LW's Harry Potter section and began confabulating that that's a good idea to flaunt in public.
oh please save us from the 'mind virus cult' mr. entropy machine god worshipper
1
1
203
What it's like sharing the evening of a one night solo business trip with Claude: "Under the applicability conditions identified by the Hotel Bochum Hbf-Süd, post-excursion handwashing faces principled contamination underdetermination rather than hygienic silence. The subject cannot allocate the felt grime between the Café Extrablatt door, the stairwell railing (green), and the red door of I/5.OG. The warm-water intervention functions as a candidate for the lenient middle: it blocks the specified bacterial substitutions conditionally, pending an admissible residual mismatch between the felt cleanliness and the actual one. The green sink is noted but not yet phenomenologically discharged."
1
2
146