Optimist. I write lots of code. Founder of five companies across education, games, VR, finance, AI companionship. I love my wife and my dog. Privacy is cool.

Australia
Since mid 2023, I’ve been working on a side-project called Neuron. It’s conceptually similar to @openclaw, but structurally very different. After using openclaw for a month, I still prefer Neuron. So we’re packaging it up for public consumption and I can’t wait to share it ASAP
1
14
469
The only thing I know about argon is that it’s inert. Are they calling their model inert?
Introducing Gemini 4 Argon – our new frontier model. It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.
1
1
76
GPT-6 Astra Ultra Fast
66
ChatGPT Images 2.5 still has the same recursive greeble issue that makes everything look like a fractal.
ChatGPT Images 2.5—faster, sharper, smarter, with better tools for creating whatever you can dream of. - Faster image generation to keep your ideas flowing - Improved fidelity for more natural, recognizable images - Consistent details across multiple edits - Comment-based edits to change only what you want
Made with AI
107
This is always my first test for a new model. Responses usually range from “I can’t help with that”, through to a nonsense response with the structure of a joke. Astra passed, but chose a safe interpretation of “explicit” IMO.
70
Noooo, anything but this. This is basically the only issue I had with Sol.
BREAKING: OpenAI just dropped GPT-6 ASTRA!!! 🚀✨ We’ve been testing it extensively at @every across coding, writing, and knowledge work. My take: it’s a big upgrade from 5.6-Sol, with some frustrating habits that keep it from matching Fable at the top end. Here’s your vibe check: - The best writing model I’ve tried. It’s fast, produces very little slop, and is easy to steer. It’s a good companion for actually working through the writing I do every day. (Not to mention, it one-shotted the first draft of its own vibe check today!) - The computer use is wild. It can go for hours at a time using complicated apps to get work done. It did the first cut of our Fable 5.1 vibe check video...kind of mindblowing - Impressive 3D games and visualizations. It can make beautiful 3D worlds from a single prompt. I one-shotted a historically accurate rendition of the Battle of Waterloo - It can overcomplicate things. (Especially at higher effort levels.) Ask for a simple interface and you get extra labels, buttons, and features everywhere. It has a habit of turning everything into a landing page. It just doesnt quite match Fable's ability to intuitively understand your prompt and do something delightful (without overcomplicating.) Net Result: If you already live in ChatGPT for Work or Codex and can afford it, it’s an easy upgrade from 5.6-Sol. The biggest proof of Astra's effectiveness at helping you do work is our vibe check. We found out it was launching at 3 AM this morning, and had a 4,000 word vibe check + video done by 2 PM. Not possible without this model. I’m reaching for Astra all day, but Fable 5.1 still gets my biggest tasks. On ambitious builds, Fable is better at understanding what I want and taking it further than I would have thought to ask. State of Play: Astra is launching to Enterprise customers today, and the rest of ChatGPT users over the coming days. Now, both OpenAI and Anthropic have a higher class of models that cost more to use. That changes who gets to use frontier AI and how. It's also a new vector of competition between them: Fable and Astra are priced at the same level. We'll see what that means for adoption in the coming days and weeks. read our full vibe check @every today: every.to/vibe-check/gpt-6-as…
1
1
124
I’ve stopped using Sol altogether this week, and I’m making a list titled “Jobs for Astra” instead.
1
82
This is the biggest issue with OpenAI models and codex right now. There’s ALWAYS a follow up cleaning step after every task with the current model.
gpt 5.6 最让我受不了的就是: 我让它做一盘番茄炒蛋,它往里还加了东坡肉。 我说有必要加东坡肉吗?它说你说得对,然后把东坡肉去掉。 我说好,你提 PR 吧。再一看,它 PR 写着「番茄炒蛋(无东坡肉)」并且注释里会写一大堆为什么本道菜不需要加东坡肉。
1
2
639
Unfortunately, GLM-5.3 thinks it’s Claude. I’m surprised that this still happens so often with open models. Why don’t they run a simple Claude replacement filter over the data they extract from the closed labs?
GLM-5.3 API is now live. - Built for coding, defensive cybersecurity, and long-horizon agentic tasks - Priced the same as GLM-5.2 - Available via the official API and partner model gateways Get started: docs.z.ai/guides/llm/glm-5.3
2
216
Grok @bot apparently doesnt have native API access to X?
1
72
“One thing worth flagging” is the worst claudism since I was absolutely right all the time.
1
106
After using opus 5 for a few days; - I like its solutions better than I like Sol’s - It is better at CAD than Sol - It does not account for edge cases at all, as opposed to Sol’s over-accounting for edge cases - I strongly dislike its personality - It is much more difficult to get it to finish a task; it is incredibly lazy and handholdy - It is much worse at general research and fact finding than Sol - Claude Code desktop app is significantly less useful and less pleasant than Codex. - Claude can natively use Codex’ Computer Use functionality somehow, that’s cool. - Having Sol work on a repo and then Opus causes a bunch of confusion for Opus due to mammoth amounts of meaningless tests that it thinks are Gospel. For example, last week I asked Sol to remove the “light/dark” label and use just the icon. This resulted in what I wanted, but also a regression test to make sure “light” or “dark” appear nowhere in the product copy.
4
230
I just bought a 3D printer, my ultimate goal is to make fun and useful robots. I’m starting from zero; I’ve been coding for 28 years, but I know very little about physical electronics. I tend to go all the way deep when I learn things, should I post when I learn what a servo is?
1
11
373
Nothing better than finding this a couple hours later, occurring 13 minutes after walking away from a running goal in codex.
1
154
I want an LLM that can disagree with me. But still do whatever the hell I asked it to do.
Imagine a LLM thats not a "Yes Man"
1
217
ChatGPT can embed interactive circuit diagrams now too!?
Wait, ChatGPT can embed sound bites now?
2
215
Kimi K3 thinks it’s Claude in its reasoning traces. This baffles me, I would think that even during distillation, they would just replace instances of Claude in the dataset with Kimi. I’m sure they could solve for the edge cases where it would spout the achievements of AI pioneer Kimi Shannon.
1
3
406
Jean-Kimi Van Damme.
1
44
It’s crazy the amount of work gpt-5.6 sol can get done correctly, while still feeling so stupid.
3
188
Wait, ChatGPT can embed sound bites now?
1
357
Stochy retweeted
Microsoft DELETED my account AND OneDrive!!?? After ACKNOWLEDGING that I’m the owner of the account and that it was compromised??? 25 fucking years of data, thousands of euros spended on games?? My son’s baby pictures? GONE! All because MICROSOFT couldn’t bring back a compromised account?? One of the biggest companies ever coulnd’t do that so they just deleted that shit like it was nothing?? Fucking shame on you!! @microsoftnl @MicrosoftHelps @MicrosoftHelpt @Microsoft #microsoft #hacked
4,588
9,526
80,595
5,842,049