SemiAnalysis retweeted
SemiAnalysis ClusterMAX Lead @JordanNanos explains how Nebius picked up demand CoreWeave couldn’t handle and then executed well enough to reach the Platinum tier of its GPU cloud rankings: “CoreWeave has set the standard. I mean, they are the best in terms of technology.” “These guys deploy 10,000 GPUs a week at the peak when they’re building data centers. That’s incredible.” “But their balance sheet is full. There’s only so many people and resources you can marshal to build out another gigawatt when they’re going to 15 and beyond.” “We have customers coming to us and saying, ‘I saw you write such good things about CoreWeave. I just have $100 million I’m trying to give to them. They’re telling me they can’t accept it until May of next year because they’re backed up.’” “And therefore Nebius has been capitalizing.” “They’ve also executed very well. The technology that they’ve developed is strong.” “We think that managed clusters are approaching this feature-complete world where it’s hard to tell the difference in quality between Nebius and CoreWeave.” “Certainly our testing experience has been really solid on both of them.”
9
15
111
29,778
The Chinese AI Infrastructure Boom: Introducing the SemiAnalysis China Datacenter Model 1,000+ facilities across 60+ operators mapped, built retail-first and flipped by AI, largest hyperscaler leases 1/5 national capacity, 100MW in 12 months, Eastern Data Western Compute newsletter.semianalysis.com/…
21
32
173
48,849
MONEY PRINTER ALERT🚨 NVIDIA vLLM B200 CAN GENERATE UP TO💰️$15 BILLION💰️OF ANNUAL PROFITS PER GIGAWATT serving the open DeepSeekv4.1 Flash model at the official interactivity & official selling prices. Using Engram DRAM offloading on NVIDIA results in a 50% increase in revenue per GigaWatt.
25
45
482
52,652
We are more bullish on China’s WFE localization after attending CSEAC 2026 in Wuxi. (1/5)🧵
10
22
274
105,031
Chinese SemiCaps are also expanding capacity faster to capture strong demand. (4/5)
2
29
12,259
AMD MI355X becomes the first official TileRT result no AgentX! This configuration achieves an astounding 470 TPS on GLM 5.3 (FP8). This is over 40% faster than GB300 TRTLLM using FP4! Great work to the TileRT x AMD team! (1/3)🧵
19
29
354
47,899
This is a disaggregated config: the prefill engine runs vLLM and the decode engine TileRT. TileRT is an ultra-low latency-focused inference engine that runs model decoding in a persistent GPU kernel to deliver faster tokens to each user. (2/3) github.com/tile-ai/tilert
1
3
21
10,750
This is just the beginning, and there is still room to improve. Early results reveal P90 TTFT is quite high compared to other engines. We expect to see this improved through optimizations to KV transfer as well as adding paged KV caching on the decode engine. (3/3)
2
17
8,343
Agentic coding puts a kind of stress on GPU clusters that providers are struggling to handle. SemiAnalysis runs it on the clusters it rates anyway, with an open-source CLI in front. "A lot of our ClusterMAX testing, we lead with our CLI, cmax. It's open source. Feel free to try it out on your managed cluster." "We do use agentic coding for testing all these clusters, and the workload is very different. The workload is very stressful on these systems." "That's why we have cmax CLI as a front end, and the agent is just a wrap around it. If there are any issues cmax CLI faces, the agent debugs into it and understands what's going on."
10
4
58
20,837
SemiAnalysis retweeted
SemiAnalysis' @JordanNanos on the neocloud talent wars: "It's across the entire supply chain. We see this in electricians, data center technicians, and operators. Salaries for electricians in Louisiana and Abilene, Texas, are up 3x to 5x." "Crusoe just came in and paid these guys so much more money. But they're out of people. They need to train. No matter what job we're talking about, people just need to be trained." "The most successful neoclouds that we're seeing have this pipeline to take talent from other related areas and get them contributing to neocloud management and day-2 operations where a software engineer can start to work as an SRE on one of these teams."
8
12
98
42,749
PSA TO NEOCLOUDS COPING ABOUT CLUSTERMAX RATING: If you run a restaurant and you get bad Google reviews, don't attack the review. Instead listen and it might make your service better. For more food-related metaphors about neoclouds, check out: newsletter.semianalysis.com/…
67
9
129
41,569
SemiAnalysis retweeted
Happy Thursday. On today's show: - @CompleteSkeptic (TypeSafe) - @JordanNanos (SemiAnalysis) - @leifthunder (Public) - @hoomanrenezhad (Solcoa) - @erika_alden_d (Pioneer Labs) - @Shalev_lif & @RomiLifshitz (Enclosure) See you on the stream.
Meta Connect Reactions, Zuck's Beer Pong Controversy, ClusterMAX 3.0, New Bentley EV nitter.net/i/broadcasts/1DGleVWwE…
1
10
44
37,051
ORACLE DELAY🚨: We already said it in our energy model on May 29 and on our twitter on Jul 17.
We're starting to see some things *ORACLE SENDS FORCE MAJEURE NOTICE OVER NEW MEXICO DATA CENTER
18
12
155
57,504