leading inference provider ops (i'm hiring!) @openrouter. opinions my own. aka tomas. 🇦🇷🇦🇷🇦🇷

vibe/acc
Pinned Tweet
the provider operations team at openrouter is absolutely cooking on automations. 727 new endpoints added in august alone, 23 endpoints a day! our systems allow 75+ providers to self serve onboard their endpoints onto our platform (gated on quality and functionality tests)
9
7
64
13,048
ragebaiting theo is a great virality strategy
9
1
112
4,584
Soft launch - if you want insights from the 130 trillion tokens that flow through OpenRouter every week, join us. Subscribe to Inferred, our data newsletter now. openrouter.ai/data
5
6
33
10,503
Live on OpenRouter 🫡
Introducing Span-01, the first hyper-parallel reasoning classifier built for unseen challenges (RLAIF). 2x cheaper, 18% better than Jev. 700x cheaper, 4% better than GPT-6 Luna. Frontier reasoning for every behavior, at classifier speed. • Span-01: #1 on Behavior Benchmark • Span-01 Lite: Better than Jev and completely free!
1
19
2,198
Toven retweeted
Introducing typesafe/jev-router: a cache-aware model router powered by Jev and @typesafeai The Jev Router picks the best model and reasoning effort for each request, balancing quality, speed, and cost. Here's how it works 👇🏻
95
136
1,787
1,095,483
Toven retweeted
Once a month --> once a week --> every single day. Keeping up with new model launches is a full time job.
25
32
509
24,406
urge to update my PFP continues to grow
4
13
766
@ItsDannnnnnnnnn yo you want another commission
1
2
134
Super fun to present at Meta Connect with @pratanchandani! Incredibly excited for what the team is cooking next.
7
3
48
1,877
OpenRouter now has tools, both free and paid, to make on demand intelligence even smarter and cheaper—starting with web search, image generation, calling other models mid-run, full shell, and even datetime().
Introducing the Server Tools Marketplace A one-stop shop for better ways to ground your agents, including web search APIs, Shell, tool search, and more 👇
4
6
77
13,962
Toven retweeted
Presenting the Memebench Grand Finals The top 4 models from round 1 face off to determine which model cooked harder on the "Why would I deceive you" meme. Which model cooked harder? A, B, C or D? 👇🏻 Vote in the poll below:
4
2
24
10,146
i’m getting 8sleep baited
1
7
646
Toven retweeted
118 trillion tokens in one week, visualized as a crowd. I made a site that turned real OpenRouter data into a tiny town where every AI lab is a shop. More tokens, bigger crowd. tt.lab.sael.net
85
130
1,839
190,805
Toven retweeted
Meet Memebench - Group B We gave 8 AI models the same meme and asked them to put @pingToven and @Audrey_Sage_ in it Second round, group B. Which model cooked harder? A, B, C or D? 👇🏻 Vote in the poll below:
9
1
27
16,481
Toven retweeted
Meet Memebench We gave 8 AI models the same meme and asked them to put @pingToven and @Audrey_Sage_ in it. First round, group A. Group B comes tomorrow. Which model cooked harder? A, B, C or D? 👇🏻 Vote in the poll below:
15
4
52
16,127
Toven retweeted
Meet Space Bunny Alpha on OpenRouter ⚡ A flash model with fast inference, adjustable reasoning, and a 1M-token context window. It takes text, image, and video input. Give it a spin and share your feedback: openrouter.ai/stealth/space-…
85
56
855
123,643
Toven retweeted
GPT-6 Sol and GPT-6 Luna from @OpenAI are live on OpenRouter! Half the price of their GPT-5.6 predecessors, Sol at $2/M input and $10/M output, Luna at $0.10/M input and $0.50/M output. On AutomationBench each one tops its predecessor's best score at a fraction of the cost per task. What each one is for 🧵
17
18
248
15,930
Live on OpenRouter :)
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
1
2
29
1,368
man i kinda miss the days where i could read every single openrouter mention on here
5
1
44
2,850
Toven retweeted
Xiaomi MiMo-V2.6 from @XiaomiMiMo is live on OpenRouter. Three new models. A 1T+ parameter flagship, an open-source MoE, and a ~10x faster variant of the flagship. All three take text, image, video, and audio with 1M context. More in the thread 🧵
42
53
1,059
70,996
Toven retweeted
The Jev moment happening now is similar in energy to the OpenClaw moment that happened in January, and also to the Opus 4.5 moment before that. Developers are scrambling to find use cases for a new hot thing, and it's a sudden blooming of creativity. Models have typically NOT optimized for specific use cases and have gone the other direction, generalizing over all of them. This could be a very important moment for the whole AI ecosystem if this turns out to be the beginning of other model skews that make the market much more diverse, such as compaction, summarization, extraction and more, all of which can now be used to optimize harnesses and decouple them from provider lock-in. We'll see what happens over the course of the next few weeks when the model labs optimize their small models more or try new ways of branding and packaging them.
18
16
225
14,905