OpenAI's Jalapeño chip beats Nvidia superchips on power and latency

Share
OpenAI's Jalapeño chip beats Nvidia superchips on power and latency

Two hardware stories today, both about who sets the price of running AI: OpenAI's first custom inference chip just posted benchmark numbers ahead of Nvidia's best, and one of autonomous trucking's quietest operators just bankrolled its expansion.

OpenAI's custom Jalapeño chip delivered 1.5–1.9x more AI work per watt than Nvidia's GB200 and GB300 superchips, with 1.7–3.6x lower end-to-end latency, according to results the company presented at the Hot Chips conference on Tuesday. The numbers come from InferenceX, an inference benchmark run by research firm SemiAnalysis, tested across GPT-OSS 120B, DeepSeek R1 and Kimi K2.5 1T. OpenAI hardware chief Richard Ho called it a "very, very significant performance advance over state of the art," noting that most systems force a trade-off between throughput and latency — Jalapeño escapes it by keeping each model's working state next to the compute serving it, instead of shuttling data across the machine mid-response.

The caveats matter as much as the headline. These are OpenAI-commissioned results on a benchmark its chip was tuned for, and the comparison ran against Blackwell-generation Nvidia systems — while Jalapeño ships only in small volumes late this year and ramps through 2027, by which point Nvidia will have moved too. Still, the direction is the news: an ASIC co-designed with Broadcom, built purely for inference, beating general-purpose GPUs on their own economics metric is the scenario Nvidia's biggest customers have been waiting for. Ho was careful to say Nvidia remains one of OpenAI's "very good partners" and that Jalapeño won't replace the whole fleet — which tells you OpenAI expects its own silicon to carry most of the inference load, not all of it.


Self-driving truck startup Gatik raised $200 million — its largest round yet — led by Qatar Investment Authority and Koch Disruptive Technologies, two months after signing a multi-year deal to run driverless Frito-Lay deliveries for PepsiCo across Texas, Arizona and Arkansas. Millennium Management, ARK Invest and Intact Private Capital joined the round, bringing Gatik's total funding to roughly $500 million since it emerged from stealth in 2019. The company says it holds $600 million in contracted revenue and operates dozens of fully driverless box trucks on commercial routes for Walmart, Kroger, Loblaw and Tyson Foods. While robotaxi operators still burn cash per ride, middle-mile freight has quietly become the corner of autonomy that actually gets paid.

What to watch: whether anyone outside OpenAI reproduces the Jalapeño numbers independently — and what Nvidia's next generation looks like before the 2027 volume ramp.

If a first-gen startup chip can nearly double the work per watt, how safe is Nvidia's pricing power? Tell us in the comments.

Read more

OpenAI's first Category 5 influence op targeted editors, not feeds

OpenAI's first Category 5 influence op targeted editors, not feeds

OpenAI banned two state-linked influence campaigns on October 8 — and the number worth sitting with is not the ban count but the rating attached to one of them: the first Category 5 operation the company has disrupted in two and a half years of publishing threat reports. The deeper signal, though, is in the fine print of what the models were actually used for. What happened OpenAI's report describes two operations it calls "false front" entities — shells that launder geopolitical messaging

The Take — OpenAI's $20B gap is definitional, and that's worse

The Take — OpenAI's $20B gap is definitional, and that's worse

The $20 billion never went missing from OpenAI's business — it was never in it. OpenAI's annualized revenue was always a number only OpenAI gets to define, and with a confidential 2027 IPO filing on record and a $1.2 trillion private round under consideration, I think a self-defined metric heading into underwriter season is worse than a number that was simply wrong. A wrong number gets corrected once; a self-defined number survives every headline it produces. Our afternoon brief on Wednesday l

OpenAI busts influence ops that planted fake stories in real media

OpenAI busts influence ops that planted fake stories in real media

The day's AI news runs through one seam: the work is showing up in places nobody planned for — inside real newsrooms, across the whole night sky, and in the M&A column. OpenAI has banned two state-backed influence operations that used ChatGPT to plant fabricated stories inside legitimate news outlets — and rated the Russian one the most disruptive it has seen in two and a half years. In a report dated October 8, OpenAI detailed "Dark Clark," run from Russia across Latin America, which ran a th

Open Source Radar — October 9: plugins, sandboxes, tokens

Open Source Radar — October 9: plugins, sandboxes, tokens

Today's open-source signal is infrastructure rather than hype: Microsoft's code sandbox reaches 1.0, Anthropic's knowledge-worker plugins keep climbing, a beloved token counter flips its default, and LocalLLaMA squeezes a usable 2B model into about 700 MB. knowledge-work-plugins (Python, ~27,900 stars, Apache-2.0) — Anthropic's repository of role-shaped plugins for Claude Cowork is the top AI repository on today's daily trending page, and the stars keep coming: roughly 2,100 more than when we