Switch language한국어
Back to the list

OpenAI Got The Whole AI Squad To Accelerate Large-Scale AI Training – AMD, NVIDIA, Intel, Microsoft & Broadcom All-In On MRC

TL;DR AI

Key summary

2 min read
  1. OpenAI and partners unveiled MRC, a new multipath networking protocol for large-scale AI training, released through the Open Compute Project.

  2. MRC splits 800 Gb/s links into multiple paths to improve resilience, cluster efficiency, and faster recovery from network failures.

  3. The protocol was developed over two years with major chip and cloud companies, including AMD, NVIDIA, Intel, Microsoft, and Broadcom.

  4. MRC is already deployed in OpenAI supercomputers using NVIDIA and Broadcom hardware, with aims to simplify scaling across very large GPU clusters.

Read the original