Inboxsmith

AI news videos

OpenAI Publishes First Jalapeno Chip Performance Results

OpenAI published the first measured performance results for Jalapeno, its first custom inference chip, on August 25, 2026. Everything below comes from that first party post.

Watch on YouTube

Transcript

OpenAI just published the first measured results for Jalapeno, its first custom inference chip. These are OpenAI's own tests.

OpenAI ran Jalapeno on InferenceX, a public benchmark from SemiAnalysis, against leading commercial systems. Across three models it reports up to one point nine times more work per watt.

Jalapeno is rated at seven hundred watts, but OpenAI measured draw at or below five hundred fifty. On the largest model tested, latency was three point four times lower.

OpenAI says its own models helped take Jalapeno from design to tapeout in nine months. Deployment inside OpenAI begins by year end. It will keep using NVIDIA accelerators too.

Inboxsmith builds AI receptionists that answer calls for small businesses. Please like and subscribe for more news.

Sources

Every claim in this video comes from the top ranking coverage of this topic. The claims and where each one came from:

We make Inboxsmith.

An AI receptionist that never misses a business call.

See how it works