The new platform combines AMD's EPYC CPUs and Instinct MI400-series GPUs in Helios rack-scale infrastructure with Cerebras' Wafer-Scale Engine solutions. This disaggregated design aims to deliver up to 5X higher tokens per second per watt by optimizing different inference stages for specialized hardware. The combined offering is scheduled to become available through Cerebras Cloud in the second half of 2026.