/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Sources: Microsoft has been rationing GPU access for teams building AI tools since late 2022; the company plans to announce Office 365 GPT-4 tools on March 16

Aaron Holmes / The Information :

The Information Aaron Holmes

Context & Ripple Effects

The arc here runs from access to allocation. In November 2021 Microsoft invited select businesses to run GPT-3 as an Azure tool, positioning itself as an open gateway to OpenAI models even as OpenAI sold its own API. By February 2023, reporting had Microsoft readying its Prometheus model integration across Word, Outlook, and other Office apps.

What changed today: sources say Microsoft has been quietly rationing GPU access among its own teams since late 2022 so that enough capacity exists to ship the Office 365 GPT-4 tools it plans to announce on March 16. The scarcest input in AI is being steered from an open Azure storefront toward Microsoft's flagship product line — and later coverage shows that squeeze extending outward, with startups unable to secure Nvidia GPUs as cloud providers divert supply to internal teams and large customers like OpenAI.

First-order effects

  • Microsoft's internal AI teams are now competing for a fixed pool of GPUs, with Office 365's GPT-4 push effectively winning priority over other product groups' experiments.
  • Azure customers and partners who assumed GPT-series models would be broadly available on demand face tighter, slower access as Microsoft reserves capacity for its own launch.

Second-order effects

  • Startups priced out of Microsoft's cloud are pushed toward rival providers — a dynamic later reporting makes explicit, with AI companies struggling to obtain Nvidia GPUs as Microsoft and other clouds divert supply inward.
  • Capacity scarcity strengthens the business case for Microsoft's in-house Athena AI chip, tested by Microsoft and OpenAI staff, reducing dependence on Nvidia's constrained supply.

Third-order effects

  • If hyperscalers systematically prioritize their own products and anchor customers like OpenAI over the open market, frontier compute allocation becomes a competitive moat rather than a commodity service — reshaping which startups can build at all.
  • Sustained rationing pushes the whole industry toward custom silicon and long-term capacity commitments, converting AI infrastructure into utility-like infrastructure controlled by a few allocators.

The trend: Cloud providers are shifting from selling open access to frontier AI models toward rationing scarce compute in favor of their own products and marquee partners.

Discussion

  • @tomwarren Tom Warren on x
    Microsoft has confirmed that “during this preview period, we are running various tests which may accelerate access to the new Bing for some users.” That's why you should be able to get access to the new GPT-4-powered Bing right away https://www.theverge.com/...
  • @aaronpholmes Aaron Holmes on x
    NEW: Microsoft is preparing to launch a suite of Office 365 AI tools tomorrow, powered by GPT-4. But Microsoft only has so many AI chips, and it's now rationing them internally while it reserves a big chunk of them for the new OpenAI-infused tools: https://www.theinformation.com/…