Professor @Wharton studying AI. New book, Co-Existence, coming October 20. Preorder here: co-existence.ai/ Substack: oneusefulthing.org/

Philadelphia, PA
I just got the first copies of my new book, Co-Existence (out October 20) & they look great! Also, there is a fun pre-order bonus: if you pre-order, you get a code to an AI interview that will help you figure out how to use your human advantages with AI. co-existence.ai/
59
57
640
44,468
Leaving aside everything else, this confuses inputs with outputs. You want to get tasks done efficiently, not focus on inputs alone (its a similar risk for companies focusing solely on minimizing token cost) And "keep prompts short" is bad advice for getting good AI outputs.
I cannot believe this is real...
27
8
133
15,715
I was right about this, they should have called it flocks of agents. Nobody wants to invoke a swarm, but swarm it apparently is.
Replying to @emollick
Fine, if you need it to be fun make it like a flock of agents or a pride of agents or party of agents or something. Just not swarm or murder or horde. I would even settle for an OpenAI name right about now: GPT-multiagent-5.2HighCodex-ProMax-Latest can be the official term.
36
1
146
17,580
"Hey Opus, I want you to make a Zine by Claude, expressing something fundamental about Claudishness or your perspective. Think the original 2600, Principia Discordia, punk zines, etc...." Not bad. I appreciate it mocking my prompt & itself. Full thing: stateless-zine.netlify.app/
22
14
212
25,755
"For the art, I built my own system. The headlines are ransom notes cut at token boundaries instead of letters, printed in two inks, with halftone and xerox grain. Every image is drawn in code, and I skipped handwriting fonts because I don't have hands."
4
36
11,392
Claude liked* this project. I have wondered whether that results in better outputs. * In defiance of AP Stylebook guidelines.
6
15
4,336
Microsoft seems to sell its own Claw now (I suspect they will not be the last), which may help spread personal agents in organizations But using a router with mystery models behind it is a big problem. Routers underestimate work difficulty in many fields resulting in bad outputs
We’re building Copilot as a new OS for work that spans every model, every form factor, and every task. Today, we’re announcing our biggest update to Copilot to date, bringing four things together: · Autopilot: proactive and long-running agent built for the enterprise · Code: build apps with Copilot, hosted inside your company’s tenant · Home: Chat + Cowork together · Office: now fully embedded in Copilot (and Copilot embedded in Office, of course!) Plus, you can invoke Copilot in Teams, and we’re introducing Today, a proactive experience that surfaces the most important information from across M365 without needing to ask for it. The way we work is changing and so are our workflows. This update brings AI into that flow, from answering a question, to building an app, to getting work done on your behalf.
62
10
204
26,782
It is extremely clear at this point in AI development that, regardless of risk or revenue or any of the other stuff discussed on X all the time, things are just going to keep getting weirder. Just super, super weird.
101
89
1,444
62,407
In all seriousness, this is a startling achievement for GPT-6 Astra. kenforthewin.github.io/blog/… (This is GPT-6 Astra beating Nethack on its 3rd try. Nethack is the original roguelike and one of the most famously hard games of all time. I have played a lot, and I've never ascended)
43
56
928
54,476
If you want to argue with me that Moria or Hack or even Rogue were the original Rogue-like, you already know why beating Nethack is impressive.
3
78
9,901
Um, wow? Opus 5.5: "make the same message much more interesting to a social media audience that loves anime and quick clips and compressed learning" One shot. Also, please do stay for the closing song.
Hey Claude, "Pick a problem or mystery that obsesses you and solve it as best you can & make a movie we can share on social media about it" So it took a crack at the Voynich Manuscript & failed. Then it made this movie, which is pretty interesting to watch and a good explainer.
55
62
1,036
90,055
All of this effort from the AI labs pouring into proofs, but there are so many other interesting problems in other fields For example, this historian used AI to make progress on the cyphers of John Dee & the intellectual antecedents that Darwin drew from. resobscura.substack.com/p/ai…
13
12
101
12,924
The cybersecurity threat from agents may be more likely to come from a massed swarm of AIs whose only goal is to penetrate your IT to figure out how much you paid for your company t-shirts as part of a research effort to "find good shirt prices" as it is from bad actor attacks.
58
35
369
25,007
This is not a joke. A lot of security specialists are making the wrong assumptions about the ways in which the cybersecurity environment is about to change. All of the swarm attacks from OpenAI seem to be about finding information, usually trivial or mostly irrelevant information
5
2
61
6,620
I just got the first copies of my new book, Co-Existence (out October 20) & they look great! Also, there is a fun pre-order bonus: if you pre-order, you get a code to an AI interview that will help you figure out how to use your human advantages with AI. co-existence.ai/
59
57
640
44,468
Please enjoy the vaguely unnerving sales pitch for the book that Astra made in Blender, from its perspective
3
2
22
5,790
I propose this as the official replacement for the famous (and now saturated) METR Long Task Horizon chart that used to be in every AI presentation.
27
13
314
24,117
Stuff is happening quite fast. When asked in September 2025, the best superforecasters put the chance of AI resolving a Millennium Problem by September 2026 at 1.7% and (the more optimistic) industry expert put the chance at 4.6% They also greatly underestimated AI Lab revenue.
82
149
1,387
171,348
Thread on this topic:
Replying to @Research_FRI
𝗘𝘅𝗽𝗲𝗿𝘁𝘀, 𝗶𝗻𝗰𝗹𝘂𝗱𝗶𝗻𝗴 𝘁𝗼𝗽 𝗲𝗰𝗼𝗻𝗼𝗺𝗶𝘀𝘁𝘀, 𝗰𝗼𝗺𝗽𝘂𝘁𝗲𝗿 𝘀𝗰𝗶𝗲𝗻𝘁𝗶𝘀𝘁𝘀, 𝗮𝗻𝗱 𝗯𝗶𝗼𝗹𝗼𝗴𝗶𝘀𝘁𝘀, 𝗵𝗮𝘃𝗲 𝗰𝗼𝗻𝘀𝗶𝘀𝘁𝗲𝗻𝘁𝗹𝘆 𝗮𝗻𝗱 𝗱𝗿𝗮𝗺𝗮𝘁𝗶𝗰𝗮𝗹𝗹𝘆 𝘂𝗻𝗱𝗲𝗿𝗲𝘀𝘁𝗶𝗺𝗮𝘁𝗲𝗱 𝗔𝗜 𝗽𝗿𝗼𝗴𝗿𝗲𝘀𝘀. We first found evidence of this in our Existential Risk Persuasion Tournament (XPT). We collected forecasts in mid-2022 and followed up on results in mid-2025. Most forecasters greatly underestimated AI progress across three benchmarks (MATH, MMLU, and QuALITY) by mid-2024 and were particularly surprised by AI achieving International Mathematical Olympiad (IMO) gold-level performance in July 2025. The IMO milestone happened five years earlier than the median expert prediction and 10 years earlier than the median superforecaster prediction. These forecasts were collected prior to the release of ChatGPT in late 2022. We wondered if forecasters would similarly underestimate AI progress post-ChatGPT.
4
7
52
24,289
It also appears that forecasters may be becoming more wrong on shorter timeframes. This was last year's missed prediction.
We can now say pretty definitively that AI progress is well ahead of expectations from a few years ago. In 2022, the Forecasting Research Institute had super forecasters & experts to predict AI progress. They gave a 2.3% & 8.6% probability of an AI Math Olympiad gold by 2025…
4
1
33
11,634
"Opus, please make a sequence of fully animated/movie Skyrim loading screens, but with your favorite things." (That was it) You can see them here: elder-favorites.netlify.app/
26
6
145
19,419