Flexible Chip Ports Could Reduce AI Networking Costs by 70-80%
Chips with flexible ports can adapt to various mesh dimensions, allowing for optimized connections within racks, to neighboring racks, and across superpods, unlike NVIDIA's fixed NVLink and PCIe categories. Utilizing fullmesh or hybrid (UB-Mesh) network topologies, as explored in the Ascend 950 paper, significantly reduces the need for switch chips and optical connectors, leading to substantial networking cost savings of 70-80%. This approach, while potentially harder to maintain, localizes high-bandwidth interactions, making it ideal for AI training workloads.