OpenAI is building Jalapeño because inference is now the electricity bill
OpenAI and Broadcom have unveiled Jalapeño, a custom LLM inference chip planned for deployment in 2026. This is not a clean break with Nvidia. It is an attempt to put cost per token under OpenAI's own control.