From IMO gold-medal-level reasoning to real-world agency. 🥇🌍
👇 Read the thread for the model and open-source links.
Built by rednote, @dotsstudioai’s dots3-note preview is an open-weight multimodal model designed for long-horizon agency in real life.
It proves its reasoning ability by moving beyond just answering questions — helping AI sustain reasoning, adapt through interaction, and execute actions across complex, ever-changing tasks.
Aug 29, 2026 · 7:03 AM UTC
15
12
14
6,398
1/ Meet dots3-note preview — an open-weight multimodal model built for long-horizon agency in real life. 🌍
Try it: studio.dots.ai/?lang=en
Api free now: openrouter.ai/dots-studio/do…
280B total / 16B active
512K context
Text + vision + audio
Created to sustain reasoning, utilize tools, and adapt through feedback over long time horizons.
Copy
1
7
98
2/ 🧠 Train agents for tasks that last hours—not minutes.
Say hello to TEMPO.
By splitting long trajectories into macro-steps, it helps agents stay on track by switching between Actor and Critic—assessing progress and course-correcting before the task concludes.
During a hidden-rule game, two distinct branches secured the exact same reward after 64 rounds.
Yet the Critic caught what the raw reward missed entirely:
Branch B misunderstood the objective and kept going in the wrong direction.
Branch A found the underlying conflict rule and moved toward a feasible solution.
Branch A: 3.8 | Branch B: 2.29
Equivalent reward. Very different progress.
1
7
18
4/ 🔁 Learn in unfamiliar environments.
Through interaction and feedback, dots3-note preview can:
Explore → Form & test hypotheses → Update memory → Reuse what it learns
This adaptive capability shines in extended interactive scenarios like ARC-AGI-3.
1
6
61
5/ Open-source the tools. Try the model.
🔎 VibeSearchBench — multi-turn search as user intent unfolds
vibebench.github.io/VibeSear…
🌎 VibeLifeBench — long-running tasks where users, plans, and conditions change over time
vibebench.github.io/VibeLife…
We are releasing two open benchmarks designed to evaluate real-life agents on extended tasks.
dots3-note preview is just the first step.
7
59














