We buy unused AI compute commitments at a discount - before they go to waste - and pass the savings on to you. Made by @keak_ai

Cheaper Inference retweeted
From prompt optimization using genetic pareto to benchmarking for field performance intel, #CheaperInference @CheaperInfer has it all if not more ! See for yourself at cheaperinference.com/signup?…
4
4
523
Live and discounted on cheaper inference
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
3
1
7
4,355
Astra is live and discounted!
18
4
102
545,457
Cheaper Inference retweeted
Cheaper Inference is up!
All the AI is down
15
7
65
17,078
Omniroute has joined cheaper inference!
We've acquired omniroute.online ! Try the #1 open source router and stop paying routing fees!
9
3
26
24,461
GPT 5.6 Luna is 60% off!
3
1
16
4,225
Time to use cheaper inference in Cursor
We’re ending our partnership with Cursor following its acquisition by SpaceX. Under our proposal, Cursor’s direct access to our models would end on November 12. We know that the people most affected by this decision are the developers who rely on OpenAI models in Cursor. We care about their experience in this transition and we’re ready to go above and beyond to support them. openai.com/index/our-decisio…
3
2
15
6,313
Cheaper Inference retweeted
We’re experiencing a surge in demand that’s pushing @CheaperInfer’s infrastructure to its limits. You may encounter some issues over the next few hours while we scale up capacity. We apologize for the disruption. We expect everything to be back to normal within approximately two hours.
7
4
29
8,910
GLM 5.3 is live and discounted 🫡
3
3
13
5,091
3.7 Flash is live and discounted by 30% 🫡
1
3
3,561
deepseek v4 pro 0813 is now live and discounted 🫡
1
1
7
22,079
Prices will come down even more over the next couple days as we get more capacity
2
2,588
This is true, DeepSeek is completely dominating our rankings
Prediction: Deepseek v4 flash will take over Claude as the no.1 winner in market share. It's the biggest story of 2026! Three reasons: 1. It's so cheap. 100x cheaper in unit token pricing. 2. It's so fast. Much much faster than Claude. My vibe check is around 2-3x faster. 3. It's so cache-efficient: most of my tokens spent are in cache read, and cache input pricing are $0.5/Mt Opus 5 v.s. $0.0028/Mt Deepseek. Deepseek has a much higher cache hit rate, which means the overall effective cost per task could be upwards of 500x cheaper. For something that's 500x cheaper AND 2-3x faster per task, why wouldn't Deepseek V4 Flash win.
1
1
10
7,110
Happened later than we thought but Opus 5 is now the top model by usage 👑
6
4,954
Opus 5 is live and discounted by 30%!
2
2
10
6,617
GLM 5.2 is now 45% off list price 🤑
2
2
2,056
Most used models today
4
6,709
Kimi K3 is available on cheaper inference and 30% cheaper than the list price!
Kimi K3 has received far more love than we expected, and our GPUs are feeling it. Over the past 48 hours, demand has pushed close to the limits of our current capacity. To protect the experience of existing subscribers, we're temporarily pausing new subscriptions and prioritizing compute for current members. Existing subscribed users are not affected. We're adding capacity as fast as we can and will reopen new subscription spots in batches. Going forward, we'll also split membership into two more focused plans: Kimi Membership for Kimi Web, App, and Work; and Kimi Code Membership for coding workflows. This will help us match compute more precisely and keep the experience stable. Thank you for your patience and understanding!
3
1,904
Kimi K3 is live and 30% off list price!
1
2
1,489
Hy3 is now free on cheaper inference
1
2
7
1,887