Fetching from the wire…
Public story · 2026-07-21 · high
The site already tops 2 gigawatts, according to an exclusive look Nadella gave podcaster Dwarkesh Patel inside the facility.
Why now: Nadella's numbers are public because Dwarkesh Patel's interview included an exclusive look inside Fairwater 2 itself.
Microsoft's Fairwater 2 site has passed 2 gigawatts, and one building there beats any AI data center now running, Satya Nadella told Dwarkesh Patel.
That 2-gigawatt figure is the real constraint behind the tight rate limits and slow inference that anyone building on frontier models runs into. The bottleneck is measured in gigawatts and construction schedules, not lines of code.
For builders, relief runs on a construction timeline. A better retry strategy doesn't fix a throttled API when inference capacity is tight. Only finished buildings like Fairwater 2 do, and that takes quarters.
Nadella didn't say when Fairwater 2 reaches full capacity, or how much of that 2 gigawatts is already serving live traffic versus still being commissioned. Patel's interview doesn't answer that either, and it matters if you're trying to plan around when the limits actually ease.
The next real relief for inference-constrained products shows up on a construction schedule, not in a model release. Watch for the next Fairwater-scale building coming online, since that's the number that resets the constraint.
Each link below shares sources, entities, or timing with this story.
In its Q3 FY2026 earnings ($82.9B revenue, +18%), Microsoft disclosed something more important than the revenue number: a structural shift in how software gets sold. The company's AI business crossed $37B annual run rate, up 123% year-over-year. But the real signal is the pric...
Amazon's AWS grew 28% to $37.6B (15-quarter high) with $200B planned capex. Google Cloud grew 63% past $20B quarterly, but executives said growth was capacity-constrained. Microsoft's AI business hit $37B annual run rate, up 123%. Meta raised AI capex guidance to $125-145B whi...
NVIDIA's Blackwell successor is in production ahead of schedule. The NVL72 rack (72 GPUs) delivers 3.6 exaFLOPS for inference, with 288GB HBM4 per GPU. NVIDIA claims 10x lower cost-per-token versus Blackwell. The Rubin CPX variant — purpose-built for million-token inference —...
On August 24 Thomson Reuters announced a frontier model it calls Thomson, built by specializing an open-source base on Westlaw, Practical Law, Checkpoint and Reuters content with input from hundreds of subject matter experts (Thomson Reuters). Stated cost: $40 million in talen...
The merge started this week as the structural precondition, bringing chat, AI coding, the Cowork research tool and new AutoPilot background agents acting across Outlook, Teams and OneDrive into one product. The rollout also retired free Deep Research within four days, which re...
On the July 29 earnings call Microsoft positioned its homegrown MAI family as the cheap alternative to its own partners, introducing "MAI thinking one" as its first reasoning model and claiming MAI Cyber One Flash, paired with a multi-agent security harness, "achieves better p...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.