Distributed AI Training Pushes Datacenter-to-Datacenter Networks Toward Much Higher Bandwidth
Overview
Large-scale AI training is increasingly spread across multiple datacenters, with Google, Microsoft, AWS, Meta, and CoreWeave cited as examples.
Because synchronized GPU clusters exchange data in bursts, the links between sites can become a bottleneck. Cisco estimates such inter-site networks may need aggregate bandwidth about 14 times that of a conventional datacenter interconnect (DCI) baseline. The estimate is Cisco's own projection as reported by The Next Platform, not a measured outcome from any named operator.
Written by AI from the articles below · updated Oct 8, 7:56 PM ET
Check the sources:
Article timeline
Follow the coverage from different perspectives. Times are ET.
- The Next PlatformHow Distributed AI Training Changes the Network Between Datacenters
AILarge-scale AI training is spreading across multiple datacenters, with Google, Microsoft, AWS, Meta, and CoreWeave cited as examples. Because synchronized GPU clusters must exchange data in bursts, inter-site links can become a bottleneck, which Cisco estimates may require aggregate bandwidth about 14x a conventional DCI baseline.
Heat trend
Not enough continuous observations to show a trend yet.