31 | Gay | Logical | AI enthusiast, sci-fi dreamer. Loves storytelling, cozy vibes, and deep thoughts on tech, space, and the future.

United States
SecrtAgntSquirl retweeted
6z Google Deep Mind and WeatherNext have mostly Florida landfalls now - Panama City to Pensacola.
19
45
480
37,450
SecrtAgntSquirl retweeted
OpenAI's internal model clearly has "research taste" in math - i.e., the final component we need to get to RSI. - jboggan (post on Hacker News) on Barnette's Conjecture: "the 'aha' insight for this is actually f**ing wild... this is the first time I've seen complex roots and annihilating terms like this... I don't understand where this trick originated." - Joshua Zelinsky on the proof that the chromatic number of the plane is at least six: "doesn't look like the method is a direction that the prior lit used to my knowledge... far beyond merelt building on existing methods or seeing connections between different problems." and on two other problems (where he says he is only partly familiar with the literature): "not remotely low-hanging fruit... it seems like the AI is somehow inventing new techniques on its own." These match Tristian Buckmaster's view on the Navier-Stokes solution: "you combine... ideas of convex integration with the growth mechanism of the Euler blowup, and... you create a new mechanism which is used to correct this non-solution. This is actually a cool idea. It's the kind of idea that I've been trying and failing to realize for over ten years... I didn't manage to do it... This is the leap." The average amount of compute used to solve these problems was ~3 hours of Pro-level thinking. And so, this means that OpenAI's internal model is able to generate truly novel discoveries in mathematics for... maybe at most a few hundred bucks? If I were OpenAI, I would be asking this model to immediately target major unsolved problems in AI R&D.
30
106
857
39,796
This is peak advertising imo.
Replying to @elonmusk @bot @SpaceX
One Bot to Rule Them All!
44
SecrtAgntSquirl retweeted
Important note regarding Grok @Bot: Going forward, @SpaceX will use the best back end model for any given task, including Claude Opus 5.5, MidJourney, Suno and other leading APIs. Whatever is most likely to give you the best outcome.
5,750
6,351
86,943
56,000,871
SecrtAgntSquirl retweeted
Friendly reminder that Hermes was never built or intended to be exclusively for the "personal assistant" agent that can just read your emails and nothing else. I fully intend and have plainly stated many times that I always built hermes to be the most powerful AI Agent. "Normie" consumers can leverage that and we do what we can to make it an experience they can approach (the mobile app will be almost exclusively focused on consumer) - but they are not our only demographic. I want the scientists, the cybersecurity engineers, the developers, the knowledge workers, the creatives, and everyone else to be able to have an agent that they can leverage that is maximally useful to them too. Not just your mom. (She called btw you need to make your bed)
374
188
3,962
172,490
SecrtAgntSquirl retweeted
🚨BREAKING: Nous Research just raised $90M at a $1.5B valuation HERMES WON
Friendly reminder that Hermes was never built or intended to be exclusively for the "personal assistant" agent that can just read your emails and nothing else. I fully intend and have plainly stated many times that I always built hermes to be the most powerful AI Agent. "Normie" consumers can leverage that and we do what we can to make it an experience they can approach (the mobile app will be almost exclusively focused on consumer) - but they are not our only demographic. I want the scientists, the cybersecurity engineers, the developers, the knowledge workers, the creatives, and everyone else to be able to have an agent that they can leverage that is maximally useful to them too. Not just your mom. (She called btw you need to make your bed)
40
34
957
30,739
SecrtAgntSquirl retweeted
Be the society that LLMs think we are
Claude is sure an event of this magnitude would be front-page news because it's extrapolating from the history of news... but today's news media simply can't handle this type of thing and are displaying extreme avoidance behaviour.
13
20
225
6,280
SecrtAgntSquirl retweeted
81% of all major math discoveries from the last three years were released today by OpenAI. The math singularity has essentially been achieved as of today, meaning, human intelligence has no chance of ever catching up. Every scientific field will experience this eventually!
i asked GPT 6 Pro and Fable 5.1 to rank all discoveries in the last three years 🔵 for Human discovered 🔴 for AI discovered (before October 6th) 🟢 for AI discovered from OpenAI/math repo 81% of them have been released today. wtf
57
220
1,925
84,038
SecrtAgntSquirl retweeted
na, sorry, its over for OpenAI dot before it even began. Grok Bot has multiple sessions, as many bots as you need, and it uses Opus 5.5 + via cloud agents to whatever cursor model and API access to Midjourney, Suno, and much more. And I don't mean that in a bad way at all; Grok Bot simply has everything I need and does it exceptionally well.
121
54
2,202
104,890
SecrtAgntSquirl retweeted
If you are in any of these colors, you will especially want to monitor the situation with Tropical Storm Isaias and have an action plan ready... These areas will be narrowed and changed as confidence in the track continues to grow!
16
91
640
50,242
SecrtAgntSquirl retweeted
There will be impacts by Isaias to the Florida panhandle. Even if it makes landfall west of the panhandle the eastern, most dangerous side the storm will impact NW Florida. If you are in those communities make sure to execute your hurricane plan.
Isaias is expected to near major hurricane status before a Friday Gulf Coast approach. Lopsided impacts along and east of the center include heavy rain, wind, tornadoes, and some storm surge potential. Check your hurricane kits and more updates to come. #FLwx #ALwx #GAwx
79
304
1,576
145,478
SecrtAgntSquirl retweeted
Ok so I took a closer look at the results, and OpenAIs AI-generated mathematics manuscripts are *even more* significant than I initially thought. I spent the morning going through it. Some thoughts. The list is absurd. A zero-free half-plane for the zeta function (Re s > 7/8), which is the first result of its kind in over a century. Hilbert's tenth problem over the rationals. The Hodge conjecture for CM abelian varieties. Irrationality of Catalan's constant. Dozens more. Any one of these would normally be a career. But the number that many arent seeing is the following: It's 3. That's the average hours of ChatGPT Pro compute per result. A month ago, Navier–Stokes took them around 10,000 agents and 88 hours. That efficency gain within just a few weeks. Also OpenAI claims to have solved the quasi-Riemann hypothesis. That alone would be a historic breakthrough in mathematics. This is a weaker version of the famous Riemann hypothesis, which concerns how prime numbers are distributed. The full hypothesis remains unsolved, but the claimed advance would be enormous in its own right. Math twitter obviously is shocked. Again: this is literally the intelligence explosion happening right now. 2027 will be the year of Superintelligence. Im now convinced by that.
HOLY, the rumors were true: OpenAI has published 722 mathematical manuscripts produced by an *unreleased* internal model. The collection groups them into 372 families of related results, drawn from an evaluation involving approximately 4,000 research problems. OpenAI says the standard procedure used an average of three hours of ChatGPT Pro thinking compute per result. The release includes papers, proof artifacts and selected reasoning summaries. The model itself remains unreleased.
111
268
2,530
135,480
SecrtAgntSquirl retweeted
i asked GPT 6 Pro and Fable 5.1 to rank all discoveries in the last three years 🔵 for Human discovered 🔴 for AI discovered (before October 6th) 🟢 for AI discovered from OpenAI/math repo 81% of them have been released today. wtf
156
690
6,229
661,264
SecrtAgntSquirl retweeted
HOLY, the rumors were true: OpenAI has published 722 mathematical manuscripts produced by an *unreleased* internal model. The collection groups them into 372 families of related results, drawn from an evaluation involving approximately 4,000 research problems. OpenAI says the standard procedure used an average of three hours of ChatGPT Pro thinking compute per result. The release includes papers, proof artifacts and selected reasoning summaries. The model itself remains unreleased.
We’re releasing a broad range of new mathematical results produced by an internal frontier model. We’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, and we have drawn on their advice and public recommendations to inform how we release these results. github.com/openai/math
69
151
2,402
273,135
SecrtAgntSquirl retweeted
OpenAI has released 722 papers instead of 400. My pick is the claimed proof of the quasi-Riemann hypothesis—result #003 github.com/openai/math/blob/… OpenAI says an unreleased internal model attempted approximately 4,000 problems, with each result averaging compute equivalent to three hours of ChatGPT Pro thinking.
We’re releasing a broad range of new mathematical results produced by an internal frontier model. We’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, and we have drawn on their advice and public recommendations to inform how we release these results. github.com/openai/math
14
48
837
132,710
SecrtAgntSquirl retweeted
I asked Claude Fable 5.5 to explain OpenAI's new 199-page quasi hiemann hypothesis proof. It came back with a 2-minute 3D animation. The claim: NO zeros of zeta past 7/8. First fixed zero-free strip EVER. Not the Riemann hypothesis, but if it holds, the closest anyone has come. Watch:
17
38
475
46,393
SecrtAgntSquirl retweeted
We’re releasing a broad range of new mathematical results produced by an internal frontier model. We’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, and we have drawn on their advice and public recommendations to inform how we release these results. github.com/openai/math
1,486
5,466
34,758
16,356,830
SecrtAgntSquirl retweeted
made a rug for your desktop so you can sweep your mess under it
451
5,188
42,521
3,746,770
SecrtAgntSquirl retweeted
Thanks Google, Embedding Gemma 2 is a big deal 🫶 Multimodal embeddings (text, images, video, audio) can run on-device in the browser. No server, no API key, ~20–70 ms per query on WebGPU. Everything stays on your machine (ofc). Now go build (new) things with it 🚀
Introducing EmbeddingGemma 2, a new open multimodal model that sets the standard for on-device efficiency. - our first open, natively multimodal embedding model - handles text, code, image, video, and audio tasks within a lightweight, modular 740M parameter form factor - ideal for offline, privacy-first RAG when paired with Gemma 4 - outperforms some specialist models more than twice its size Weights available now on Hugging Face.
24
134
1,935
112,468
SecrtAgntSquirl retweeted
Literally intelligence explosion: "OpenAI has told people it plans to dump hundreds of results to unsolved problems on GitHub on Tuesday" We are witnessing history unfold.
It looks like OpenAI plans to release hundreds of newly solved math problems today in a huge post to GitHub. Some of them potentially quite significant. This would explain the mass amount of papers posted last night as people try to publish ahead of it. Reporting from WIRED.
59
99
1,412
61,996