General Tech
Business Insider1 day ago
0

AMD takes a shot at Nvidia by betting on AI's next big shift

AI

AMD is partnering with Cerebras to push disaggregated inference, splitting AI workloads across specialized hardware, and claims its new Helios server system outperforms Nvidia's Vera Rubin NVL72 in cost-efficiency.

AMD takes a shot at Nvidia by betting on AI's next big shift

Intelligence Insights

Context + impact, normalized for TechCulture.

The Big Picture
AMD CEO Lisa Su announced a partnership with chip startup Cerebras to advance disaggregated inference, an approach that splits AI workloads across different hardware types rather than relying on a single chip. AMD argues that processing prompts and generating responses are fundamentally different tasks, so Helios handles high-volume requests while Cerebras' wafer-sized chip specializes in rapid response generation. The company claims Helios delivers up to 30% more inference tokens per dollar than Nvidia's Vera Rubin NVL72 rack. This shift aligns with industry trends noted by UBS, which sees disaggregated inference as a response to current architectural limitations, though orchestration challenges remain. Major AI labs and cloud giants including OpenAI, Meta, and Microsoft already use AMD infrastructure, and the partnership will bring Helios into Cerebras data centers later this year.
Why It Matters
AMD's partnership with Cerebras signals a strategic bet that AI inference will shift from single-chip solutions to disaggregated architectures, challenging Nvidia's dominance. By splitting workloads across specialized hardware, this approach could lower costs and improve efficiency for AI deployments, but it also introduces new orchestration challenges that will test the industry's ability to integrate diverse chips seamlessly.

Deepen your understanding

Use our AI to break down complex signals.

Select an AI action to generate more depth.

AMD president and CEO Lisa Su at London Tech Week, smiling in front of a black background.
AMD president and CEO Lisa Su at London Tech Week, smiling in front of a black background.
AMD president and CEO Lisa Su.

Leon Neal/Getty Images

  • AMD is betting the future of AI inference won't rely on a single chip.
  • A new Cerebras pact reflects a shift toward "disaggregated inference."
  • AMD says its new Helios server system rivals Nvidia in performance and cost.

AMD is making a bet about the future of AI: one chip shouldn't rule them all.

CEO Lisa Su announced Thursday that AMD is teaming up with chip startup Cerebras on a new approach to AI inference, which is the process of generating responses from AI models.

Increasingly, chipmakers are pursuing "disaggregated inference," which splits workloads across different types of hardware. AMD's partnership with Cerebras shows the company is betting big on this approach.

Traditionally, the same hardware handled both processing a prompt and generating an answer. AMD argues those are fundamentally different jobs. Helios, its latest server system, is designed to process huge volumes of requests, whereas Cerebras' giant, wafer-sized chip specializes in generating near-instantaneous responses.

The partnership will bring Helios into Cerebras' data centers later this year.

Demand for chips from companies like AMD, Nvidia, and Broadcom has skyrocketed in the AI boom. Nvidia dominates chip design for AI training, and the competition has intensified as AI companies shift focus from training models to putting them to work.

The AMD and Cerebras pact aligns with a broader shift that analysts say is already underway, with UBS writing in June that the limitations of current architectures "are driving a shift toward disaggregated inference."

UBS wrote that Nvidia — through its integration of AI hardware startup Groq — and Amazon Web Services are also pursuing similar setups to improve efficiency and lower costs. That said, UBS wrote that disaggregated inference presents new challenges around "orchestration" — or getting different chips to work together seamlessly.

At Advancing AI, AMD unveiled Helios, its latest server system that bundles several types of AI chips, which is its answer to Nvidia's Vera Rubin NVL72 rack. AI labs and cloud giants using AMD's infrastructure include OpenAI, Meta, Microsoft, Oracle, and Anthropic, with which AMD announced a multibillion-dollar infrastructure partnership on Wednesday.

AMD also used the event to take direct aim at Nvidia, claiming that Helios delivers up to 30% more inference tokens per dollar than Nvidia's Vera Rubin NVL72 rack.

"Every Helios can deliver more performance for the largest models, more capacity for longer context, and the bandwidth to scale across thousands of racks," Su said Thursday at AMD's Advancing AI event.

Have a tip? Contact this reporter via email at gweiss@businessinsider.com or Signal at @geoffweiss.25. Use a personal email address, a nonwork WiFi network, and a nonwork device; here's our guide to sharing information securely.

Read the original article on Business Insider
Startups Hardware Big Tech AI

Intelligence Exchange

0

Log in to participate in the exchange.

Sign In

Syncing Discussions...

Finding Related Intelligence...
AMD takes a shot at Nvidia by betting on AI's next big shift | TechCulture