A gap between Microsoft's AI chip claims and what's actually installed

Started by Hawk, Today at 08:08 AM

Previous topic - Next topic

0 Members and 1 Guest are viewing this topic.

Topic: A gap between Microsoft's AI chip claims and what's actually installed   Views(Read 42 times)
Active members in this topic:
Hawk(1)

Hawk

A Guardian investigation has found an apparent discrepancy between what Microsoft has publicly said about its AI capacity and the number of advanced AI chips actually in operation. Microsoft reportedly targeted having 1.8 million AI chips installed globally by the end of 2024. Nearly two years later, in the middle of a $280 billion expansion, internal documents seen by the Guardian show the company has 2.2 million AI chips installed, well below what some outside experts had assumed given the scale of the spending.

Sources within Microsoft told the Guardian the company's total AI chip count has barely moved over the past year. Some of the gap may trace back to Microsoft's partnership with OpenAI, since the exact commercial terms aren't public and that unit's deployments may not appear in the documents the Guardian reviewed. There's also a question of how much of Microsoft's headline capacity is actually operational, with the company's flagship US project, a pair of datacentres in Wisconsin and Georgia called Fairwater, still apparently far from fully live despite CEO Satya Nadella saying in April that the Wisconsin site was going live.

Academic scrutiny adds another layer here. Shaolei Ren, a professor at UC Riverside, told the Guardian that Microsoft's own sustainability reports, which separately disclose electricity usage and are audited by a third party, suggest the company's actual 2024 AI capacity was closer to 1.2 gigawatts, a figure that would imply roughly 4 million chips if Microsoft genuinely added 5 gigawatts of AI datacentre capacity over the past two years. Ren's point isn't that Microsoft is lying exactly, more that the company is giving insufficient context about what its own capacity claims actually mean.

This lines up with a broader shift in how Microsoft talks about its own bottleneck. Nadella has said publicly on a podcast that the company is no longer chip supply constrained, framing the real limit now as a lack of ready to use power and building shells to actually plug hardware into, not chip availability itself, which is a notable reversal from a company that flagged GPU availability as an investor risk factor in earlier annual reports.

So the honest picture seems to be a genuinely messy mix of measurement ambiguity. Partnership structures that obscure the real numbers, and datacentre projects that are announced as live well before they're actually running at full capacity, rather than a single clean explanation either way
Just here for the craic :)

Save money on everyday spending Free cashback on thousands of retailers
View offer