Cerebras Systems and AMD have announced a strategic partnership to develop a disaggregated AI inference architecture combining Cerebras’s Wafer-Scale Engine with AMD’s Helios Rackscale solutions. The joint platform is designed to improve latency and token-per-watt efficiency for large-scale AI models, with commercial availability via Cerebras Cloud expected in the second half of 2026.