Steven Gonsalvez

Software Engineer

OpenAI Jalapeño: Better than Nvidia Blackwell

Why CEREBRO kept it

OpenAI custom inference chip, LLM optimization

The text below is an automated extraction of the article at https://newsletter.semianalysis.com/p/openai-jalapeno-better-than-nvidia, stored verbatim in the public cerebro-vault repository. Copyright remains with the original publisher (newsletter.semianalysis.com).

OpenAI has spent the past couple years quietly building “Jalapeño,” an inference chip just announced at Hot Chips. Rumors of a successful tapeout had been swirling for a while. But now we have details. OpenAI invited us to look at their chip, go to their labs to check out how real it is, and benchmark it with our InferenceX suite. In June, OpenAI unveiled the chip program in partnership with Broadcom, built from a blank slate exclusively for LLM inference. Design work began in the middle of 2024, going from initial team hiring to manufacturing tape-out in ~16 months, an extremely fast ASIC dev

Community take

Inference won't democratize, it'll entrench oligopoly as only scale players hit unit economics.

Backlinks

Appeared in 1 briefing

Related

Shares tags: ai/llm-mechanics · cerebro/signal · release-notes

Also from newsletter.semianalysis.com

Only signal from newsletter.semianalysis.com so far.