ex-google brain and deepmind. phd in neural nonsense from stanford.

San Francisco, CA
Pinned Tweet
Real-world models are here! Stoked to share how we're bringing real-world locations to life by integrating Street View into Genie. Try it now at labs.google/fx/projectgenie and read the blog for more info: blog.google/innovation-and-a…
18
94
625
229,082
Ben Poole retweeted
I've been on a Blender kick with Opus 5.5. Its better 3D modeling and vision mean you can build an entire world from a single prompt. Historically accurate San Francisco Market street in 1906, pre-earthquake
91
47
1,119
145,866
Ben Poole retweeted
I just realized Opus 5.5 can use Blender to make claymations with just a single prompt in claude dot ai now
102
106
2,620
257,385
sketch-to-simulation with Opus 5.5
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
67
128
2,347
251,204
❤️🚲💑
Got married in the park (in a bicycle parade) last weekend and @peterhartlaub put together such a lovely article about it! sfchronicle.com/totalsf/arti…
12
4,212
Ben Poole retweeted
We are hiring research scientists and engineers🧑‍🔬 on the Gemini Omni team! Both in the US (Mountain View / SF) and Europe (London / Zürich). Apply here: google.com/about/careers/app…
13
68
781
154,715
Ben Poole retweeted
Atlas is live! A spatial foundation model, trained in-house from scratch. After shipping RTFM last year, I was convinced a single camera-conditioned model could unify most spatial generation tasks. Atlas is that bet at scale. The core idea👇
Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. Model the world, move the camera, and simulate space & time.
4
16
177
61,648
Ben Poole retweeted
Two years later our first video model - FLUX 3 Video - is finally released🥹. This is such a strong, fun and versatile model made by the best team in the world @bfl_ai. We will add further variants and generation from references in the near future. I am also super excited about the upcoming open weight, action prediction and image variants of FLUX 3.
38
51
500
25,976
Ben Poole retweeted
Fortunate to work with this team every day, ACT-2 is a big leap. Working on real Solves is a blast. Next Stop: Recursive Solve Improvement
Introducing ACT-2 Preview The first robotics model to unify broad generalization with high reliability. A single fine-tuning example can teach Memo a new behavior that generalizes. Zero shot, real unseen homes, 99% success rate.
6
5
48
3,172
Ben Poole retweeted
1/2 Introducing Nano Banana 2 Lite! We are working hard at the next frontier model, but in the meantime enjoy this fast and cheap Nano Banana (Pico Banana?). It returns images in only a few seconds, and with most of the smarts of the NB2 model! deepmind.google/models/gemin…
2
8
47
4,192
Ben Poole retweeted
We’re shipping 2 major releases:
 🔘 Nano Banana 2 Lite: our fastest and cheapest Gemini Image model 🔘 Gemini Omni Flash: now available via the Gemini API and in @GoogleAIStudio to help developers generate and edit high-quality videos.
101
197
1,123
282,034
After 15 years (4 internships and exactly 8 years full-time), today is my last day at Google. Growing up as a scientist at Brain and DeepMind was an incredible privilege, and I'm so thankful for the people who made it such a special place. I went from editing decision trees by hand as an intern, to exploring representation learning, latent-variable models, and diffusion as an AI researcher, to kicking off a new wave in generative 3D with DreamFusion, to scaling up generative models with Veo, Genie, and Omni. We're just beginning to build systems that can understand and simulate the real world, and I'm excited to see what’s next 🧠🐸🚀
113
30
1,480
121,503
Thanks to @Cannes_Lions for recognizing Project Genie with the Grand Prix in Digital Craft: "The real breakthrough happens when creativity unlocks what technology can become" canneslions.com/news/cannes-…
2
1
35
4,821
Huge congratulations to the Project Genie team on taking home the Cannes Lions Grand Prix for AI Craft! 🎉 Thank you @Cannes_Lions for recognizing how Genie pushes beyond conventional creative limitations to achieve outcomes unattainable without frontier AI… Another unexpected consequence of scaling world models :) Try it yourself here: labs.google/projectgenie
25
60
325
112,920
Ben Poole retweeted
Exciting news: Gemini Omni Flash is now #1 in the Video Arena (both Text-to-Video and Image-to-Video)! For Text-to-Video this is a massive +158 pt improvement over Veo 3.1 (1080p) and a large +61 pt lead over the next best model, Seedance 2.0. Congrats @GoogleDeepMind for this huge milestone!
We’re dropping Gemini Omni: our first step towards a model that can create anything from anything - starting with video. It combines Gemini’s intelligence with our generative media systems - representing a leap forward in world understanding, multimodality, and editing 🧵
145
244
1,517
433,392
Ben Poole retweeted
🌍 Project Genie access is expanding even more! Starting today, Google AI Ultra 5X subscribers (our latest tier!) globally can access Project Genie. Try it out here! labs.google/projectgenie
We love seeing all of the worlds you’ve been creating with Genie. SO much so, that we’re excited to announce Project Genie is now fully available to all Google AI Ultra subscribers globally (18+).
33
118
750
936,965
Ben Poole retweeted
Diffusion Circle happening today @CVPR #CVPR2026 Upper Lobby F at 3:30pm, come join us!
1
17
53
22,000
Ben Poole retweeted
Ideogram 4.0 has four key components: (1) a frozen Qwen3-VL-8B text encoder (2) a 34-layer single-stream DiT (3) a flow-matching Euler sampler with asymmetric CFG (4) a frozen FLUX.2 VAE We thank the open-source community and give back by releasing this with open weights.
2
6
64
8,199
Ben Poole retweeted
Try out our image to video feature in Gemini Omni Flash! Interested to see what people are able to do.
Bring your images to life ⚡️ Upload your picture as a first frame and add a prompt to generate your own unique video with Gemini Omni Flash. Tag us and share your results!
1
3
12
3,579
Ben Poole retweeted
Diffusion Circle with forest vibes! 🌲 Since the original Diffusion Circle organizer @sedielem cannot make it to CVPR in person this year, we coordinated with him and are happy to host the Diffusion Circle this time. Stop by if you are interested in flow maps, multimodal diffusion models, variational flow matching, action prediction, and all the other hot denoising stuff. @sumith1896 and @dustin_podell will coordinate. Meeting point: Upper Lobby F / Exhibit Hall F Entrance; see attached pics. Happening Saturday, June 6, 3:30pm!
8
16
106
16,287