Custom AI inference silicon: why OpenAI built Jalapeño and the ASIC-vs-GPU cost-per-token war
OpenAI and Broadcom's Jalapeño ASIC targets roughly 50% lower cost per token. Five hyperscalers now run their own inference silicon, and the GPU premium is fading. What the custom-silicon war means for investors.