News report

OpenAI Pairs Jalapeño ASIC Racks With AMD EPYC Turin Hosts

OpenAI's Jalapeño deployment uses dual-EPYC Turin host trays with 1.5TB of DRAM, prioritizing platform maturity while Vera remains under evaluation.

On this page
  1. Jalapeño's production system uses AMD EPYC Turin host trays
  2. OpenAI says the Turin choice was about reducing deployment risk
  3. The host choice fills in a missing part of OpenAI's custom-silicon story
  4. Jalapeño remains an inference accelerator, not a replacement for every GPU workload

Jalapeño's production system uses AMD EPYC Turin host trays

OpenAI's first-generation Jalapeño inference ASIC is being deployed in a rack-scale system whose host side uses AMD EPYC Turin-class processors. SemiAnalysis documents a paired layout: a host rack contains 16 Katsu CPU trays, each corresponding to one of 16 Vindaloo accelerator trays in the adjacent ASIC rack.

Each Katsu host tray carries two Turin-class AMD EPYC CPUs and 1.5TB of DRAM, plus two E1.S and two M.2 SSDs. The host and accelerator trays connect through external PCIe links. The important point is architectural rather than a consumer-CPU win: Jalapeño is an accelerator, while the EPYC systems provide the mature host-compute layer around it.

OpenAI Jalapeño host and accelerator rack layout
LayerReported configurationRole
Host tray2× AMD EPYC Turin-class CPUs, 1.5TB DRAMCPU host for a corresponding accelerator tray
Host rack16 Katsu traysOne host tray per Vindaloo accelerator tray
Accelerator tray8 Jalapeño ASICsLLM inference compute
Accelerator rack16 Vindaloo trays128 Jalapeño ASICs per rack
Host-to-accelerator linkExternal PCIe connectionsSystem I/O between paired trays

OpenAI says the Turin choice was about reducing deployment risk

OpenAI hardware VP Richard Ho told Tom's Hardware that selecting Turin was a pragmatic decision intended to de-risk the Jalapeño system and move quickly. He said the Turin platform did what OpenAI needed and that its partners already had experience with it.

Ho contrasted that maturity with NVIDIA's standalone Vera CPU, describing Vera as being somewhat behind on maturity for this particular design decision. That should not be read as a general performance verdict on Vera versus EPYC: OpenAI is also among organizations exploring Vera, and the comments concern the host choice for the current Jalapeño deployment.

The host choice fills in a missing part of OpenAI's custom-silicon story

OpenAI and Broadcom unveiled Jalapeño in June as OpenAI's first Intelligence Processor, built specifically for LLM inference and intended for multi-generation, gigawatt-scale deployment. OpenAI's announcement established the accelerator program but did not make the CPU host the headline.

The newly detailed rack architecture shows that custom AI silicon does not eliminate the general-purpose host layer. OpenAI is combining its purpose-built inference accelerator with an established x86 server platform while the custom accelerator, networking and software stack mature together.

Jalapeño remains an inference accelerator, not a replacement for every GPU workload

OpenAI positions Jalapeño around LLM inference. SemiAnalysis has published detailed architecture and performance analysis after observing the system and benchmarking in OpenAI's lab, but it also notes important evidence boundaries: some benchmark inputs and numbers came from OpenAI, and the full preferred AgentX suite had not been run.

That distinction matters when interpreting the rack design. The EPYC host decision is concrete deployment information; it does not establish that x86 is universally superior for AI hosts, that Vera will not be used in later OpenAI systems, or that Jalapeño replaces NVIDIA and AMD accelerators across training and other workloads.

Sources

Primary and technical sources

These sources support the reporting and analysis above. Current stories are updated when later evidence materially changes the facts.

  1. 01 OpenAI

    OpenAI and Broadcom unveil LLM-optimized inference chip
  2. 02 SemiAnalysis

    OpenAI Jalapeño: Better Than Nvidia Blackwell
  3. 03 Tom's Hardware

    OpenAI's Jalapeño ASICs are deployed alongside AMD EPYC Turin CPUs as hosts

Related

Technical guide

16 GB vs 32 GB vs 64 GB RAM for Gaming PCs

Choose 16 GB, 32 GB, or 64 GB of system RAM for a gaming PC by measuring the games and simultaneous workloads you actually run instead of relying on a universal capacity rule.