The official AI benchmark of the vibe coding movement @bridgemindai

United States
Bridgebench retweeted
GPT 6 Astra made this video entirely with code. Same prompt that I used for the Claude Opus 5.5 video. Anthropic is so far ahead right now.
Claude Opus 5.5 made this video entirely with code.
34
10
270
30,610
A lot of people are saying Anthropic has already nerfed Claude Opus 5.5. We have every BridgeBench result for Opus 5.5 from day 1 of launch. We're retesting soon. If the scores drop, you'll see it here first. Have you noticed a difference?
225
84
4,001
232,940
Bridgebench retweeted
A week ago you could hit your Claude Code limit in 30 minutes with Fable 5.1. GPT 6 Astra in Codex could burn your whole week in 4 hours. Today I ran Opus 5.5 for 8 hours and didn’t hit my limits a single time. We are accelerating.
68
30
1,405
40,270
Bridgebench retweeted
Claude Opus 5.5 just ONE SHOT this marketing video. Everything that you see was built entirely with code. This is getting crazy....
30
5
247
13,769
Bridgebench retweeted
Claude Opus 5.5 made this video entirely with code.
93
58
1,128
72,149
Grok 4.7 is MORE EXPENSIVE than GPT 6 Astra. Grok 4.7 may be the worst model release ever. There is no reason that anyone should be using Grok.
65
25
792
40,324
GPT 6 Luna is absurdly good for the price.
How did OpenAI make GPT 6 Luna this good? GPT 6 Luna: under a penny Claude Fable 5.1: $1.67 Claude Opus 5.5: $0.90 This is the cheapest model in the world and it's good.
6
1
99
16,007
Claude Opus 5.5 is the #1 front end model on BridgeBench. 1. Claude Opus 5.5: 950 2. Claude Fable 5.1: 860 2. Claude Opus 5: 860 4. Claude Fable 5: 800 5. Kimi K3: 780 Anthropic holds the top 4 spots. Not a single OpenAI model in the top 7. Nobody is close to Claude on UI design right now.
16
7
249
11,917
Bridgebench retweeted
Claude Opus 5.5 completely changes what is possible. Work that used to take me 2 weeks now takes me 2 hours. I still can't wrap my head around that. It's Fable 5.1 level intelligence at $4 in and $20 out, it's faster than Opus 5, and the limits finally let me use it all day. No model has ever given me all three at once. The world is changing so fast and most people have no idea yet.
50
35
899
27,284
OpenAI models are way more token efficient than Claude. Opus 5.5 is the most token hungry model on the chart. Over 4x more than GPT 6 Astra for the same task. More tokens means higher costs, slower outputs, and burning through your usage limits faster.
89
6
402
37,654
Bridgebench retweeted
Opus 5.5 is having an Opus 4.5 moment.
106
38
2,019
153,738
Bridgebench retweeted
I love how token efficient OpenAI models are. Tasks that take Claude Opus 5.5 over an hour are completed by GPT 6 Astra in 15 minutes or less. While Claude Opus 5.5 is the better model overall I am finding myself using GPT 6 Astra and GPT 6 Sol for a lot of tasks that need to be done quickly and efficiently.
75
11
696
37,211
MiMo V2.6 Pro is a good model. But it's way too slow. Same prompt. 3 models. Rocket launch. GPT-6 Astra: $0.81, 4m 45s Claude Opus 5.5: $1.52, 12m 3s MiMo V2.6 Pro: $0.06, 30m 50s MiMo is 13x cheaper than Astra but took over 6x longer. Cheap doesn't matter much when you're waiting 30 minutes for one output.
33
9
271
24,975
Bridgebench retweeted
Grok 4.7 might be the worst model I have ever used. Stupid. Slow. Token inefficient. Expensive. It somehow got worse than Grok 4.6 in every way that mattered. Grok 4.6 was fast and cheap. That was the whole pitch. Grok 4.7 lost both and gained nothing. I cancelled my SuperGrok Heavy subscription. There is zero reason to use Grok models right now. I spent all summer saying scaling laws would save Grok. They did not. Waiting for Grok 5.
91
27
641
45,417
Claude Opus 5.5 vs GPT 6 Sol on the black hole merger test. Opus 5.5: $0.69, 6m 8s GPT 6 Sol: $0.08, 1m 33s Opus 5.5 is a significant leap for design. It's not close. But Sol did it at 1/8th the cost and 4x faster. Which one would you use?
47
7
232
30,536
Bridgebench retweeted
GPT 6 Sol beats Fable 5 on BridgeBench. At 1/5 of the cost. GPT 6 Sol: 643 Fable 5: 628 GPT 6 Sol is one of the best models I have ever used. And on a ChatGPT Pro subscription the usage limits are basically unlimited.
106
25
774
54,936
We gave 4 models the same prompt: build a sunset ocean. Claude Fable 5.1: $1.67, 6m 20s Claude Opus 5.5: $0.90, 6m 45s GPT-6 Astra: $0.59, 3m 54s GPT-6 Sol: $0.07, 44s Which one did the best?
78
11
474
61,892
Bridgebench retweeted
Claude Code limits are INSANELY good now. Dozens of Claude Opus 5.5 agents running at once. After 2 hours I am at 94% of my session. Fable 5.1 used to burn through it in 30 minutes. 4x the limits on a better model. This is the first time Claude Max has actually felt like 20x.
92
37
1,261
51,112
Bridgebench retweeted
Opus 5.5 reminds me of Opus 4.5 and that is the highest compliment I can give a model. Opus 4.5 was the moment Anthropic gave us frontier intelligence at a price you could actually run all day. Opus 5.5 is that moment again. Fable 5.1 level performance at $4 in and $20 out. It is faster too. I have been the loudest voice complaining about Claude limits for a month. They just raised them with this release. Credit where it is due. I genuinely do not know if I will use Fable 5.1 anymore. Opus 5.5 does the same work at 40% of the price. This model changes the frontier. Anthropic cooked.
25
10
475
17,177