Chief Research Officer at @OpenAI. Coach for the USA IOI Team.

Pinned Tweet
We solved Navier-Stokes! Many of my colleagues left their fields because they believed that working on AI would be the fastest way to solve their fields' grand challenges. It's surreal to see the first of those challenges fall and to watch a future we believed in become real.
We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
464
389
5,484
999,300
It’s hard to overstate @michpokrass's impact on our consumer products. ChatGPT has had a rebirth under her leadership, and a lot of what we've built has been driven by her vision and conviction. Really grateful to work with her!
just crossed four years at openai! the special thing about this place is the constant capacity for rebirth. for all its faults, there is nowhere quite like it. the team makes high conviction, contrarian bets over and over, and they are mostly right. it's a new company every three months. i have really enjoyed my part in some of the bets, and i'm excited to share some of the current ones we are cooking up now. onward!
4
2
414
34,615
Mark Chen retweeted
Not to forget our friends across the pond.... Next week we'll host a second smaller get-together to celebrate GPT-6 in London! Our team there delivered core contributions to Astra, and we'd love to celebrate with you :) events.openai.com/gpt-6-lond… (click my link not Sam's)
We want to celebrate with people using GPT-6. We did this for GPT-5.5 and it was really fun. We’re getting together in SF on September 16 to talk about the model, what we should build next, and mostly just to hang out. Apply by Sep 10: gpt6-launch-event.openai.cha…
7
8
144
24,501
The AMS’s statement is short and sweet: ams.org/news?news_id=7686 "The news today of progress on resolving the Navier–Stokes problem, one of mathematics' great longstanding challenges concerning the equations that govern the flow of fluids, represents a milestone advance in human knowledge. This story began with Navier, Stokes, Leray, and Ladyzhenskaya and has culminated in the recent breakthroughs of Córdoba and Martínez-Zoroa, then — assisted by new technologies — Alpöge and Buckmaster, with the final steps taken by OpenAI mathematicians. The purpose of mathematics is human understanding, and this achievement, and the process that led to it, will bear fruit for a long time to come."   Ravi Vakil, President of the AMS John Meier, CEO of the AMS
12
100
737
78,497
Two things to distinguish: Did any human or agent look at user data as part of the Navier Stokes effort? No. Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company.
“we cannot rule out that de-identified data derived from their usage of our products helped improve our models.” i mean props to them for straight coming clean. (so far the proof looks more along the lines of another euler blowup proof we had, off of whose ansatz naming we were making really stupid puns like “smooth criminale”, unlike the much better “ideal fluids explode”, Tristan) so i’ll now give a bit on my thinking here. i actually woulda been pumped to collaborate on this, there are a lot of people at oai i like (ok, clearly some were indirectly dicks to me because of being part of the whole situation, but im a big boy, i still like them), idgaf about authorship on that step anyway, coulda been me Tristan and every fte at oai for all i care (on that Tristan would disagree:p). but on hearing the loud convo in the hallway, especially the part where a millennium prize was offered if i’d just be removed from the paper, it was kinda clear the die had been cast and things were locked. pretty wacky, unstrategic, and unnecessary, since on my side things were mostly me and claude having a good time yoloing random stuff in the corner rather than anything institutional. i also like the idea of the labs cooperating, and even better on scientific progress. it’s a shame!
556
120
2,280
667,724
Mark Chen retweeted
This Is How Air Actually Moves Through Your Home
62
1,662
11,179
3,373,171
Agree with @JensenHuang: we’re entering the AGI era. The AGI era must also be the alignment era. We need to teach AI to love humanity and train AI monitors as capable as the AIs they supervise. Said best by @merettm in this thoughtful, sobering piece: openai.com/index/an-alien-mi…
GPT-6 Astra, trained on ~100K+ NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years. AGI has arrived. Congratulations @OpenAI team. 400K GPUs coming online next.
117
125
1,486
164,480
Mark Chen retweeted
Alongside the launch of GPT-6 Astra, we're committing $1 billion to subsidize Daybreak access and frontier capabilities to help frontline defenders protect essential services and critical infrastructure. Welcome to the AGI era for cybersecurity. openai.com/index/daybreak-fo…
70
83
1,097
177,927
Mark Chen retweeted
My biggest wow moment using GPT 6 Astra had nothing to do with the visual posts you've been seeing... Don't get me wrong, those are awesome and fun... But the biggest wow moment was when I gave Astra context over everything about my business and then gave it the prompt below. Email, Slack, Business Texts, Notion, and every meeting note... everything I've been doing (Which involves like 20 different projects with many teams over the past few months)... Here's the prompt I used which was not really well thought out just a wispr flow ramble... All I can say is the response to this prompt was the most useful thing an AI has ever given me... ""I'm about to plan the next month, quarter, and year" Make one document that aligns me on all of those timeframes. Look for my flaws, look for my strengths, and highlight those. Look at ALL of my activities, which activities do you think i'm wasting the most time. Which activities should I do more of? Of the people I work with who seem the most dependable? Where am I the least dependable? If I could only upskill in one area over the next year what would it be and why? Where can I organize my company better, and how would I do that? Use text when necessary, use charts/visuals when necessary, but never ever fill space for the sake of it. Every graphic and word should matter. I want to know about finances, relationships, business model, everything. ""
71
90
2,224
276,871
Mark Chen retweeted
A concerningly common take seems to be that keeping Chain of Thought monitorable doesn't matter because interpretability will save us, or it's already useless This is total bullshit. CoT is our best current tool for safety & interpretability, losing it would be a major tragedy
40
127
1,390
61,166
Mark Chen retweeted
GPT-6 Astra recreated the Palace of Fine arts in Blender. This is favorite building in San Francisco because it was built for the World's Fair in 1914 where the steam locomotive and telephone were showed off. It was a time when technology gave us all a deep sense of optimism for the future. I feel like some of that sense has been lost since, but I hope models like astra can help restore it, and push us towards a more hopeful future.
226
412
5,704
2,613,150
GPT-6 Astra is here! This is a big moment for our research team - years of work on pretraining, reinforcement learning, and post-training have come together in our most capable and aligned model yet. It can build and test software, work across apps on your computer, and even help you take a crack at open scientific problems! Capabilities that felt like grand challenges a few years ago have become tools people can actually use. One example is Computer Use - if you’ve tried this before and felt like it was too slow or not good enough, I encourage you to give it another shot. We’ve come a long way since Operator, and it “just works” now. We’re also asking these systems to act on your behalf for more consequential work. Agents needs to stay aligned with your goals and values, think transparently, and respond to oversight even when tasks become difficult. We’ve made substantial progress on these behaviors in Astra, alongside stronger monitoring that can stop potentially unauthorized actions. That work is part of what makes this release possible. I think alignment is one of the most important research frontiers in AI, and it remains far from solved. Our ability to understand and align models has to keep pace with model capabilities. We want to give people more room to think, build, and discover with increasingly powerful tools that remain *under their control*. Huge thanks to the researchers and teams who got us here. There’s a lot more work ahead, and I’m incredibly excited about what we can make possible in the near future!
This is GPT-6 Astra. Anything you can do on a computer, Astra can do for you. Fast.
113
164
2,490
620,848
Mark Chen retweeted
We temporarily slowed some frontier training to strengthen security and monitoring. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations help us test safeguards and gather more evidence of alignment. I expect confidence in safety to increasingly set the pace of AI development. We urgently need tools for labs and countries to coordinate on this, which is why I signed Pacing the Frontier. In the meantime, we’re taking practical steps ourselves - and will continue to share what we learn as our approach evolves. openai.com/index/pacing-mode… pacingthefrontier.com/
94
99
1,374
397,307
Many of our strongest researchers are choosing to focus on alignment, but we're also hiring! If you want to work at a frontier lab which takes alignment seriously and doesn't pretend it's solved, please apply.
These are incredibly misleading headlines – @OpenAI Preparedness is very much alive and well by any meaningful definition Our subteam – RSI/misalignment Preparedness – is doing more urgent work than ever, and has never been more empowered to do so!
15
24
397
53,509
Mark Chen retweeted
WOW!! 🤯 Among many jaw-dropping results, this proves NP-hardness of the Closest Vector and Nearest Codeword Problems for *polynomial* approximation factors, for the first time ever, and via a totally new approach (Reed-Solomon techniques). Amazing!
yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra proofs, complete with lean certificates and CoT walkthroughs for each of them. The results are wide-ranging, from von Neumann algebras (disproof of Connes' Rigidity Conjecture) to better bounds for high dimensional sphere packing, for circuit complexity, for monochromatic triangles in multicolored graphs, and more. More thoughts here: openai.com/index/ten-advance…
11
65
657
170,332
yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra proofs, complete with lean certificates and CoT walkthroughs for each of them. The results are wide-ranging, from von Neumann algebras (disproof of Connes' Rigidity Conjecture) to better bounds for high dimensional sphere packing, for circuit complexity, for monochromatic triangles in multicolored graphs, and more. More thoughts here: openai.com/index/ten-advance…
277
944
6,675
4,279,697
Another step towards intelligence too cheap to meter!
major price cuts today: *80% drop for GPT-5.6 Luna, now $0.20 per million input tokens and $1.20 per million output *20% drop for GPT-5.6 Terra, to $2/$12 *GPT-5.6 Sol gets Fast mode in the API, up to 2.5x the speed for 2x the price, same intelligence
14
8
269
20,458
This one is personal for me: 4 years ago I decided to go into AI development because I thought there was a chance that by the end of the decade AI could be better than me at math: well it happened ahead of schedule, 2026 instead of 2030 😅. It is 100% clear that science is being fundamentally transformed before our eyes, and it is NOT aspirational anymore to say that AI will accelerate science. But this will only happen if scientists can actually use SOTA models. Science is best done by scientists. The new ChatGPT for Academics is designed exactly for this purpose: we want to empower our users and not hold this power for ourselves. I can't wait to see where the frontier of knowledge will be very soon, we have so many questions we want answers to!!! piped.video/MLehRytu9Zo
46
56
448
67,970
I’m incredibly excited to welcome Jacob Tsimerman to OpenAI! His mathematical talent is clearly extraordinary, but so is the seriousness and depth with which he engages on AI safety.
Jacob Tsimerman, who won a fields medal, perhaps math's most prestigious prize, just announced at a press conference that he is "pivoting toward AI safety" and will be going to OpenAI soon
20
54
1,000
118,495