/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI partners with Microsoft, AMD, Broadcom, Nvidia, and Intel researchers to detail the Multipath Reliable Connection (MRC) protocol to help scale compute

OpenAI is getting creative to deal with the industry's imminent compute crunch. … The protocol, which has been in the works for two years …

The Deep View Nat Rubio-Licht

Context & Ripple Effects

OpenAI’s compute strategy has long been closely tied to Microsoft infrastructure: earlier coverage described Azure as its primary cloud venue and a dedicated Azure supercomputer for distributed-model workloads. More recent coverage also placed OpenAI alongside Meta, Microsoft, and Google in work on Triton, aimed at making AI code run efficiently across chips beyond Nvidia’s CUDA ecosystem.

MRC extends that arc from compute procurement and chip software into the network layer. The notable feature is the participation of researchers tied to Microsoft, AMD, Broadcom, Nvidia, and Intel, spanning cloud, accelerators, networking, and processors.

First-order effects

  • OpenAI and its infrastructure partners now have a jointly detailed protocol aimed at making large distributed compute systems scale more reliably across multiple network paths.
  • Microsoft and the participating hardware vendors gain a common technical reference point for testing or supporting MRC in the AI infrastructure stacks they build around OpenAI-scale workloads.

Second-order effects

  • A protocol designed to improve distributed connectivity could reduce the extent to which scaling gains depend solely on adding larger accelerator clusters, increasing attention on networking software and hardware as AI-system bottlenecks.
  • The cross-vendor effort complements work such as Triton: together, such software-layer initiatives can make it easier for operators to combine hardware from multiple suppliers rather than optimize exclusively around one vendor’s proprietary stack.

Third-order effects

  • If adoption broadens beyond the participating researchers, AI infrastructure competition may shift further toward open or interoperable systems software that coordinates heterogeneous chips, servers, and network equipment.
  • That outcome is not assured: MRC’s significance will depend on implementation support and real-world deployment, but the collaboration signals that reliable interconnect behavior is becoming a strategic part of AI compute scaling.

The trend: AI builders are increasingly treating networking and portable systems software—not just accelerator supply—as core levers for expanding large-scale model training and inference capacity.

Discussion

  • @openai @openai on x
    MRC is already deployed across all of OpenAI's largest supercomputers that we use to train frontier models, including our site with @Oracle Cloud Infrastructure (OCI) in Abilene, Texas, and in @Microsoft's Fairwater supercomputers. MRC is now available through the [video]
  • @openai @openai on x
    We've partnered with @AMD, @Broadcom, @Intel, @Microsoft, and @NVIDIA, to release Multipath Reliable Connection (MRC), a new open networking protocol that helps large AI training clusters run faster and more reliably, with less wasted GPU time. https://openai.com/...
  • @sk7037 Sachin Katti on x
    Today we shared MRC ( https://openai.com/...), a networking protocol developed with @Microsoft, @nvidia, @AMD, @Broadcom, and @intel to improve how large AI training systems move data and recover from failures. This innovation has come full circle for me personally, it was
  • @amd @amd on x
    At AI scale, raw bandwidth breaks down. What matters is resilience, recovery, and consistency under load. AMD in collaboration with Microsoft and @OpenAI defines a proven solution approach to AI networking with MRC. Learn more now: https://www.amd.com/... [image]
  • @nvidiadc @nvidiadc on x
    Gigascale AI needs networking built for scale, resilience and openness. NVIDIA Spectrum-X Ethernet now supports MRC, a new RDMA-based protocol that improves throughput, availability and failure recovery for large-scale AI training. Used by @OpenAI, @Microsoft, and @Oracle. Now [i…
  • Yang Zhou Yang Zhou on linkedin
    Just saw OpenAI release their Multipath Reliable Connection (MRC) protocol: https://lnkd.in/...  My feeling is that their techniques are basically …
  • Sachin Katti Sachin Katti on linkedin
    Today we shared MRC (https://lnkd.in/... a networking protocol developed with Microsoft, NVIDIA, AMD, Broadcom, and Intel to improve how large AI training systems move data and recover from failures. …
  • Nat Rubio-Licht Nat Rubio-Licht on linkedin
    Exclusive from me this morning: OpenAI and some of the industry's biggest names have come together to fix two of the biggest issues in networking: Congestion and failure. …