Jeffrey Ling retweeted
Took a little less than 3 years but can finally say we did zero to one. All credit goes to the incredible people I get to work with. Never would I have imagined getting to work with such a world-class team on such a history-defining technology.
Cognition has crossed $1B in annualized revenue run rate. This milestone belongs to our customers. Here's how a few of them are building with Devin.
16
8
308
14,549
Jeffrey Ling retweeted
Cognition has crossed $1B in annualized revenue run rate. This milestone belongs to our customers. Here's how a few of them are building with Devin.
156
160
1,839
822,651
Jeffrey Ling retweeted
Demand for SWE-2 in Desktop & CLI is unprecedented. Glad to announce, we were able to secure additional compute. We're making SWE-2 free in Devin Cloud for Pro, Max & Teams subscribers - until October 8. Enjoy!
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.
96
76
766
165,672
We're going beyond single agents to Devin swarms. Shout out to @dkrajews123 @ido_pesok for the ship!
Some engineering tasks require reasoning across the codebase: Which code is safe to delete? What queries are slowing performance? Introducing Code Scans: codebase-wide audits for any goal. Devin investigates, reports findings, and opens the PRs. Powered by Agentic MapReduce.
1
1
11
473
Jeffrey Ling retweeted
GPT-6 Astra helps @cognition’s Devin back up “it works” with tests before the team ships.
63
52
816
93,988
Jeffrey Ling retweeted
think I unlocked godmode in @DevinAI… the best part? it barely touches usage here's why Fusion is one of my favorite AI updates of the last week after testing: → after 5 hrs of using Fable 5.1 High across four repos as lead agent and SWE-2 medium as sidekick… I'm only at 15% of my weekly Devin Max quota lol → comparing it to my Max plans on every other AI platform… this setup stretches my plan the furthest after hundreds of tests (for me) → on the Artificial Analysis coding agent index, Fable 5.1 + SWE-2 in Fusion scored 61.7 at $7.90 a run… while Fable 5.1 solo in Claude Code scored 62.2 at $12.36 → with Astra it's 39% cheaper but 2.7 points lower… Astra + SWE-2 in Fusion scored 58.9 at $4.54 a run vs 61.6 at $7.47 for Astra solo in Codex → setup wise: if you wanna get started, install Devin CLI or the desktop app, select Fusion, pick Fable 5.1 (high) as lead and SWE-2 (medium) as sidekick with fast off (you can steal my exact setup) with Fable, that's less than a point difference on the index, yet 36% cheaper per run. lowkey Cognition cooked here bros.
Introducing Fusion in Devin CLI The most efficient frontier harness for Fable & Astra; 39% cheaper across coding benchmarks. Pick your favorite model for planning and a cost-effective model for execution.
35
12
205
21,868
Jeffrey Ling retweeted
Devin is the first cloud agent with macOS, Windows, and Linux . Nobody else has it because it’s really fckin hard to do. 🥲 We rebuilt storage, networking, VNC, and CUA from scratch - in Rust! Super proud to be working alongside the great minds of @cognition and @dioxuslabs
Special delivery: Devin just got a Mac 🍎 Now Devin can: 1. Build & test apps on its own Mac VM with iOS simulator 2. Send a screen recording via Slack 3. Send a TestFlight link so you can start using it 📲
33
32
522
60,858
Jeffrey Ling retweeted
Special delivery: Devin just got a Mac 🍎 Now Devin can: 1. Build & test apps on its own Mac VM with iOS simulator 2. Send a screen recording via Slack 3. Send a TestFlight link so you can start using it 📲
188
304
3,287
2,020,251
Jeffrey Ling retweeted
We ran Devin Fusion on our Code Migration benchmark and the performance of Devin Fusion with GPT-6 Astra and SWE-2 sidekick places it on the Pareto Frontier of the benchmark. *Code Migration Bench is our benchmark that determines if models can reimplement working programs in another language.
Replying to @cognition
We partnered with @ArtificialAnlys and @ValsAI to evaluate Fusion across several coding agent benchmarks. The cost savings hold across the board while maintaining frontier performance.
9
7
125
18,517
Jeffrey Ling retweeted
Fusion has been on a task for two hours and my weekly usage hasn't budged. What the heck I'm not used to this I really need a /side or /btw so I know if it is making good progress or not
9
4
56
9,825
We independently benchmarked Devin Fusion for its release today - this is the first time a multi-model coding agent has been included on the Artificial Analysis Coding Agent Index, and it effectively retains Claude Fable 5.1 and GPT-6 Astra performance while reducing costs Devin Fusion runs a frontier lead model with a cost-efficient sidekick. We tested configurations from Cognition combining frontier models from Anthropic and OpenAI with their new SWE-2 (medium) as a sidekick model. Configured with Claude Fable 5.1 (xhigh) + SWE-2 (medium), Devin Fusion scores 62 on the Coding Agent Index v1.5, while with GPT-6 Astra (xhigh) + SWE-2 (medium) it scores 59. The Fable configuration has the higher score, while the Astra configuration is 43% less expensive and completes tasks 31% faster. Congratulations to @cognition on the release! See below for our results and analysis 🧵
95
89
1,508
2,168,196
Jeffrey Ling retweeted
A big improvement for CLI fans who want to use more Astra or Fable. We've tuned Fusion so that the delegation is unnoticeable day-to-day. Worth giving it a try. Hats off to the team that made this possible @joon_h_lee @spdling @mapeiyuan @morgantepell @maaslalani
Introducing Fusion in Devin CLI The most efficient frontier harness for Fable & Astra; 39% cheaper across coding benchmarks. Pick your favorite model for planning and a cost-effective model for execution.
4
3
73
4,475
RT @joon_h_lee: Fusion is now available in Devin CLI and Desktop! You can pair Fable or Astra with sidekicks like SWE-2 (which is free for…
1
1
We added Fusion to Devin CLI! Fusion combines two models in one harness for frontier performance at lower cost. It was fun to tune the delegation patterns for different models. Turns out Fable and Astra have very different behaviors when delegating work to the sidekick -- Fable does more exploration up front, while Astra is more rigorous in its code review. Try different pairings and see what works best for your tasks!
Introducing Fusion in Devin CLI The most efficient frontier harness for Fable & Astra; 39% cheaper across coding benchmarks. Pick your favorite model for planning and a cost-effective model for execution.
12
6
119
7,154
Jeffrey Ling retweeted
Introducing Devin Voice 🦦 ☎️ Your favorite AI software engineer just got a landline. You say it, Devin ships it. Powered by GPT-Live and our new SWE-2 model.
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.
118
130
1,582
333,249
Jeffrey Ling retweeted
Welcome @jkelleyrtp and the entire Dioxus team to Cognition! Dioxus is one of the most beloved open source Rust frameworks. We're proud to continue supporting Dioxus, Blitz, Taffy, and Subsecond while bringing the team's expertise to Devin's VM, computer use, and testing.
We have some very big news to share. 🎉 Today, Dioxus Labs is joining @Cognition to accelerate the development of Devin, Cognition’s autonomous cloud coding agent. We are excited to continue building Dioxus while also helping Cognition shape the future of software engineering.
20
24
295
44,317
Jeffrey Ling retweeted
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.
438
508
6,755
2,265,096
Jeffrey Ling retweeted
here's the writeup of the factorization! cognition.com/blog/factoring…
Last week we published a factorization of RSA-260. Today, we’re sharing the methodology of how Devin and a Cognition researcher built the world’s fastest GPU optimized lattice siever, to make factoring numbers 10x cheaper than the previous state of the art: cognition.com/blog/factoring…
15
65
695
233,265
Jeffrey Ling retweeted
Last week we published a factorization of RSA-260. Today, we’re sharing the methodology of how Devin and a Cognition researcher built the world’s fastest GPU optimized lattice siever, to make factoring numbers 10x cheaper than the previous state of the art: cognition.com/blog/factoring…
26
98
779
209,195
And we're just getting started
The world needs far more software than it can build. Cognition exists to change that. We’ve just raised over $2B at a $48B valuation, led by a16z, Accel, Founders Fund, General Catalyst, and Avenir. Since our round in May, run-rate revenue has grown from $492 M to almost $900 M.
34
2,076