Trending
Hong Kong property firm ITC inks memorandum to build 1GW data center near Shanghai Volta Data Centres appoints Jason Liggins as CEO Pantheon Atlas secures grid approval for 1GW Croatia data center Developer eyes 930-acre data center in Ransom Township, Pennsylvania AMD and Schneider Electric launch reference standards for Helios AI rack DOE: AI Data Centers Are Transforming America’s Transmission Map AI Is Redefining Data Center Ownership: Multiple Assets, Multiple Timelines The real challenge for liquid cooling isn’t deployment – it’s scale PJM grid hit by voltage disturbance after data center load abruptly drops offline – report IBM acquires HRL Laboratories to further boost quantum computing R&D efforts Nokia networking wins bolster latest earnings Google announced as end user of 8 million sq ft data center in Columbia, Georgia Sponsored: AI’s impact on data center infrastructure – is this the dawn of “the flux capacitor”? Eurus Energy & Toyota break ground on wind-powered data center in Hokkaido, Japan Sponsored: The key to the data center power problem

AMD partners with big chip co. Cerebras for ultra-low-latency and high throughput AI inference system

AMD has entered into a technical partnership with rival chip company Cerebras.. The two will develop a disaggregated AI inference solution combining AMD’s new Helios rackscale solution with the Cerebras Wafer-Scale Engine, which features the company’s large AI chip.. The Cerebras Wafer-Scale Engine – Cerebras.

The combined offering is set to be integrated in a single inference workflow, with AMD Helios providing a high-performance, scalable throughput engine for ultra-high throughput, processing prompts and large context windows.. Cerebras will provide ultra-fast, ultra-low latency memory-bandwidth-intensive token generation..

Cerebras plans to deploy AMD Helios systems in its own data centers, with the joint solution expected to become available initially through Cerebras Cloud in the second half of 2026. It will be available more broadly following the cloud rollout.. The increasing demand for ultra-low latency token generation was the reasoning behind Nvidia’s semi-acquisition of Groq, with the company set to deploy the LPX rack featuring Groq LPUs.

Intel has similarly partnered with SambaNova.. “At Cerebras we build the world’s largest and fastest chip,” CEO Andrew Feldman said, detailing a number of large customers. “They deploy us because we’re blisteringly fast.”. He added: “AI has moved from being a novelty to being useful, and in some domains, a necessity.

When it’s a necessity, people want to use it quickly. We saw a partnership where we could extend our footprint in ultra-low latency… It’s really something amazing.”. 16 Feb 2026

 

Join the conversation

Your email address will not be published. Required fields are marked *