OpenAI and Broadcom unveil Jalapeño, a data-center chip for scalable LLM inference

1 min read
Source: Ars Technica
OpenAI and Broadcom unveil Jalapeño, a data-center chip for scalable LLM inference
Photo: Ars Technica
TL;DR

OpenAI and Broadcom introduced Jalapeño, a purpose-built ASIC designed from scratch for large-language-model inference in data centers, with early testing claiming substantially better performance per watt; development took nine months and is part of a broader effort to own more of the AI stack and reduce reliance on Nvidia, with deployments planned by year-end as the silicon race heats up.

Share this article

Want the full story? Read the original reporting

Read on Ars Technica