tech

OpenAI and Broadcom unveil LLM-optimized inference chip

OpenAI and Broadcom (NASDAQ: AVGO) today unveiled Jalapeño, OpenAI’s first Intelligence Processor: an accelerator architected around OpenAI’s vision for the future of LLM inference, and the first AI accelerator in a multi-generation compute platform the companies are building together to make advanced AI faster, more reliable, and more accessible to more people.

OpenAI and Broadcom unveil LLM-optimized inference chip

TL;DR

  • OpenAI and Broadcom have unveiled Jalapeño, OpenAI's first Intelligence Processor designed for LLM inference.
  • Jalapeño is part of a multi-generation compute platform aimed at making advanced AI faster, more reliable, and accessible.
  • The chip was designed from scratch based on OpenAI's understanding of LLM fundamentals and future roadmap.
  • Early testing suggests Jalapeño will offer substantial performance per watt improvements over current technologies.
  • The architecture is optimized to reduce data movement and balance compute, memory, and networking resources.
  • This collaboration represents a commitment to scaling the physical infrastructure for AI over the next decade.
  • Jalapeño is designed for both current and future LLMs across the industry, suitable for interactive LLM products at scale.
  • OpenAI is designing more of its infrastructure stack, from chip architecture to product experience, to optimize for efficiency and access.
  • The development cycle for Jalapeño, from design to manufacturing, took nine months, leveraging AI models to accelerate the process.
  • The goal is to democratize AI by making advanced models dependable and affordable for a wider range of users.