AMD and Cerebras announced on Thursday they are teaming up to offer a combined platform for faster AI inference through Cerebras' cloud service. The system is expected to become available in the second half of 2026.
Under the partnership, the two will develop a disaggregated AI inference solution combining AMD's new Helios rack-scale product with the Cerebras Wafer-Scale Engine for disaggregated AI inference, expected to deliver up to five times higher tokens per second per watt compared to a Cerebras Wafer-Scale Engine-only configuration.
"AMD Helios delivers leadership performance and scale for the broadest range of inference workloads. Together with Cerebras, we are extending that leadership into the most latency-sensitive applications and creating a powerful new platform for real-time agentic AI," AMD CEO Lisa Su commented.
https://breakingthenews.net/Article/AMD-Cerebras-collab-on-AI-inference/66763569
No comments:
Post a Comment
Note: Only a member of this blog may post a comment.