AMD and Cerebras are collaborating on solutions for ultra-low-latency inference processing, with initial offerings expected later this year. These solutions will be available in the Cerebras Cloud and targeted at large customers.
AMD is deepening its integration into specialized AI inference hardware, expanding its market beyond general-purpose GPUs into dedicated AI accelerators.