AMD Partners with Rival Cerebras on AI Server Rack

· Source: The Information · Field: Technology & Digital — Artificial Intelligence & Machine Learning, Cloud Computing & IT Infrastructure · Depth: Intermediate, quick

Summary

AMD announced a strategic partnership with Cerebras, a rival AI server chip developer, to integrate AMD's server racks with Cerebras wafers. This collaboration aims to enable the simultaneous operation of both companies' chips, specifically for AI workloads. The initiative represents a practical application of "disaggregated inference," a technique designed to run the same AI model across different chips to optimize processing and manage complex workloads more efficiently. This move allows for combining specialized hardware from distinct vendors, potentially enhancing overall AI server capabilities and offering more flexible deployment options for demanding AI applications, marking a significant step in heterogeneous computing for AI.

Key takeaway

For AI Architects evaluating server infrastructure, this partnership signals a shift towards heterogeneous computing. You should consider solutions that integrate diverse chip architectures, like AMD and Cerebras, to optimize performance and cost for specific AI workloads. Explore disaggregated inference strategies to maximize resource utilization and flexibility in your deployments.

Key insights

AMD and Cerebras partner for disaggregated inference, running AI across rival chips simultaneously.

Principles

Method

The article describes connecting AMD server racks to Cerebras wafers to run AI across both companies' chips simultaneously.

In practice

Topics

Best for: CTO, VP of Engineering/Data, Director of AI/ML, AI Architect, AI Hardware Engineer, MLOps Engineer

Related on AIssential

Open in AIssential →

Editorial summary, takeaway, and curation by AIssential. Original article published by The Information.