The Superintelligence Cloud

San Francisco, CA
Based in United States
Pinned Tweet
AI infrastructure is the largest industrial buildout of our lifetime. Lambda is assembling the leadership team to match the opportunity ahead. Today, Lambda welcomes global infrastructure operator Michel Combes as CEO and former AT&T Communications CEO John Donovan as Chairman of the Board. Co-founder Stephen Balaban takes on the CTO role full-time, shaping the technology that will define the next decade of AI compute. Read the exclusive from @business: bloomberg.com/news/articles/…
8
8
57
12,761
The organizations developing AI have the deepest understanding of their technology, and we are pleased to support companies like @nvidia that are taking a leadership position to facilitate responsible use and development. The NVIDIA Open Agent Safety Platform is a concrete step toward setting safer boundaries for AI agents, enabling teams across industries to run frontier training and inference safely.
Today, with over 100 industry partners, we introduced the NVIDIA Open Agent Safety Platform, bringing together OpenShell and Sentry. Artificial intelligence is extraordinary technology that will advance discovery, productivity, security, health, and prosperity for generations to come. But its full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility. This is bigger than a single product. It's the beginning of an open ecosystem to build the trust layer for safe agent systems. Together, we are building the foundation of the AI economy. Trust and innovation are not in conflict. Safety is how trust is earned. We must build not only the most capable AI, but the most trusted AI, so that this extraordinary technology can realize its enormous promise for the world. nvda.ws/4hcoq7m
11
1,176
Lambda is coming to Mayes County, Oklahoma. The new data center will deploy closed-loop cooling to reduce water use, and Lambda will pay 100% of its energy costs. It's expected to generate $500M in tax revenue over the next decade and create up to 1,000 construction jobs: lambda.ai/blog/new-lambda-da… We’re excited to join the Mayes County community and look forward to listening, learning, and working alongside local leaders and residents. We’re committed to being a good neighbor and long-term partner as this project moves forward. Read more in @tulsaworld: tulsaworld.com/news/state-re…
7
11
40
5,355
AI compute is not becoming a commodity. The industry is building vertically integrated infrastructure, reaching from software and networking to the physical sites and power that make large-scale compute possible. The stakes are physical. Lambda Co-Founder & CTO Stephen Balaban makes that case on the Main Stage at Yotta 2026. He will also discuss our target of 3 GW of AI compute capacity by 2030.
2
4
17
2,357
The argument in one line. AI infrastructure is moving from traditional data centers to gigawatt-scale campuses, and the companies building it are changing with it.
1
252
Day 2 raises a different question. If you are building an AI-native enterprise, when do you rent compute, when do you reserve dedicated capacity, and when does owning it become the competitive advantage? Lambda President of Cloud Services David Ward joins a panel on exactly that. It takes place Wednesday, September 30, 11:40 a.m.–12:25 p.m., in the Future of Compute track. Panel page at yotta-event.com/the-ai-nativ…
1
1
180
Stephen Balaban (@stephenbalaban) has been through 5 pivots in 14 years. He started with facial recognition and AI filters before building the workstation and GPU compute business. The interview is a look at the long, non-linear path behind Lambda, and the decisions that shaped it along the way. Listen to the full conversation with @LambdaAPI’s CTO below on the @FoundersInArms podcast with @immad and @rajatsuri. foundersinarms.substack.com/…
1
6
29
2,625
Ask a 3D vision-language model what's near the table and in front of the curtain, and it might guess "sewing machine." The right answer is a tray rack. CVP (UC San Diego + Lambda, WACV 2026) fixes this with a target-affinity token for task-relevant objects and an allocentric grid for global context. Against Video-3D-LLM: • SQA3D EM: 58.6 → 62.3 • Scan2Cap CIDEr: 83.8 → 90.5 • Better on all 5 benchmarks tested Full results across ScanQA, SQA3D, ScanRefer, Multi3DRefer, and Scan2Cap, plus how the central/peripheral split works: lambda.ai/blog/cvp-spatial-r…
7
20
2,282
.@stephenbalaban on why Lambda prices compute like a utility, dollars per GPU, not per token, in Georgia Butler's latest Tokenomics piece for @dcdnews: "…in the same way that a utility provider might look at selling dollars per kilowatt hour, we look at dollars per GPU."
DCD Magazine issue 62 out now: The coming wave dlvr.it/TVWlln
1
6
1,137
AI for molecular dynamics has a data problem. The trajectories you need to train on are expensive. EGInterpolator (ICLR 2026, with Stanford) learns molecular structure first from abundant conformer data, then uses scarce MD data to learn motion. On the DRUGS benchmark, it reduced the gap to reference simulations by 73% for bond angles, 78% for bond lengths, and 24% for torsional motion. The structure-first ablation also matters. Removing pretraining increased mean JSD from 0.173 to 0.332 for bond angles and from 0.142 to 0.386 for bond lengths.
3
3
3
915
Lambda retweeted
Wearable AR has had a decade of attempts and still hasn't found the product. The pattern is familiar: a technology looks inevitable long before it's actually usable. Worth remembering the next time something seems like it should obviously work. More from our conversation with @stephenbalaban of Lambda on Founders in Arms. Link in bio.
12
3
28
17,638
Lambda retweeted
More GPUs do not guarantee more progress. Better orchestration does. With @lambdaAPI, we increased GPU utilization from ~20% to 43% and cut queue starvation by 74%. Make every GPU count. #SPREEAI #Lambda #AIInfrastructure
GPU utilization increased from ~20% to 43% on a reservation of 96 NVIDIA H100 GPUs, while cutting queue starvation by 74%, with blocked jobs falling from roughly 10 per day to around 4. That’s the concrete result @SpreeAI saw after fixing their orchestration. When a unified diffusion model requires 80–100 GB of memory, you can’t simply throw workloads at a cluster and expect to use those GPUs efficiently. SPREEAI was dealing with workload fragmentation, ad-hoc submissions, and storage I/O blocking that left expensive GPUs idle. Working with Lambda’s ML engineering team, they implemented MLflow-based experiment orchestration with structured queuing and workload matching. They also connected Lambda’s Prometheus APIs to Grafana for real-time visibility into utilization gaps. The video testimonial covers how they diagnosed the bottlenecks and what the remediation looked like.
1
1
1
521
GPU utilization increased from ~20% to 43% on a reservation of 96 NVIDIA H100 GPUs, while cutting queue starvation by 74%, with blocked jobs falling from roughly 10 per day to around 4. That’s the concrete result @SpreeAI saw after fixing their orchestration. When a unified diffusion model requires 80–100 GB of memory, you can’t simply throw workloads at a cluster and expect to use those GPUs efficiently. SPREEAI was dealing with workload fragmentation, ad-hoc submissions, and storage I/O blocking that left expensive GPUs idle. Working with Lambda’s ML engineering team, they implemented MLflow-based experiment orchestration with structured queuing and workload matching. They also connected Lambda’s Prometheus APIs to Grafana for real-time visibility into utilization gaps. The video testimonial covers how they diagnosed the bottlenecks and what the remediation looked like.
1
15
1,852
Huh, this is rather fascinating. Lambda Labs, the AI neocloud company, originally sold an HD camera-enabled baseball cap. The idea, as I surmise, was to acquire images for convolutional neural networks, a different technology from the now more well known large language models.
People ask us, how did you know cloud compute would be so important back in 2015? And I say I didn’t. I don’t know shit. 1517 invested in a hat company.
1
4
625
Evaluating image editing models with a single score hides the nuance behind a lower score. EdiVal-Agent (ICLR 2026) splits evaluation across instruction following, content consistency, and visual quality. Its agentic judge reached 81.3% agreement with human judgments, compared with 75.2% for a VLM-only evaluator and 68.9% for CLIP_dir. lambda.ai/blog/edival-agent-…
1
7
7
1,041
The failure that stands out is how quickly multi-turn editing exposes weak preservation. Each new instruction has to land without undoing earlier edits or changing unrelated content. Errors compound even when single-turn results look strong.
1
4
479
EdiVal-IF measures instruction following. EdiVal-CC covers content consistency, while EdiVal-VQ checks visual quality. The paper came from a collaboration between UT Austin and UCLA, with Microsoft and Lambda. arxiv.org/abs/2509.13399
233
Lambda retweeted
!!!
12
2
48
2,245
Build or buy AI? Wrong question. Lambda's @boborado is on the main stage at AI Infra Summit 2026 with @JLL CTO Yao Morin, @usbank EVP & Chief AI Officer Prashant Mehrotra, and @carrier Chief Data & AI Officer Arun Nandi, and the room's landing on the same answer: it's build AND buy. New term coined: valuemaxxing. Using value per task as a lens to govern resource allocation, while architecting around agility (easy model swaps) and availability (SLAs and model lifecycle).
1
3
14
1,096