OpenAI Details Jalapeño Inference Chip Test Results
The company plans to deploy the custom ASIC in its data centers by year end.
OpenAI released early test results for Jalapeño, its first custom inference chip. The company stated it will begin deploying the chip in its own compute infrastructure by year end as the first step in a multigenerational roadmap, with Gen 2 already in development. The presentation occurred at the Hot Chips conference and used open source models as the benchmark focus. Multiple independent analysts posted observations on the efficiency comparisons shown in the shared materials.
No MTP, No PD disaggregation, Pure TP, still beats NVIDIA's Vera Rubin NVL72 on a third-party model, with A0 stepping. And B0 is 25% better. NVIDIA GPUs become HBM wrappers.
