Enterprise DNA

Omni by Enterprise DNA

Enterprise DNA Resources

Latest AI and industry news. Practical AI operating-system thinking for owners, operators, and teams doing real work.

220k+

Data professionals

Omni

AI agents and apps

Audit

Map the manual work

News AI News

OpenAI's Jalapeño inference chip beats Nvidia's Blackwell in first published benchmarks.

SemiAnalysis's independent InferenceX tests (GPT-OSS 120B, DeepSeek R1 670B, Kimi K2.5 1T) show OpenAI and Broadcom's 700W ASIC delivering 1.5-1.9x.

Enterprise DNA |
OpenAI's Jalapeño inference chip beats Nvidia's Blackwell in first published benchmarks.

AI Pulse · Frontier Labs Watch

The play

Keep inference vendors flexible, and request workload-specific benchmarks before committing to custom silicon or long-term capacity contracts.

OpenAI’s custom inference chip, Jalapeño, has posted its first published benchmark results against Nvidia’s GB300 system. In independent InferenceX tests across three large models, including GPT-OSS 120B, DeepSeek R1 670B and Kimi K2.5 1T, the 700W chip delivered 1.5 to 1.9 times more throughput per kilowatt. It also showed up to 3.6 times lower latency than Nvidia’s 1,400W GB300, according to the published report.

Plain English, OpenAI may be able to run certain AI workloads faster while using materially less power. That matters because inference, the cost of running models in production, is where AI bills build up. Training a model is expensive, but serving thousands or millions of requests every day is where power, hardware capacity and response time become operational issues.

There are important limits. Jalapeño is only targeting low-volume production in late 2026, and it has not been tested against Nvidia’s upcoming Vera Rubin generation. These results don’t mean Nvidia has lost its lead. They do show OpenAI’s chip effort is looking like a credible performance option, rather than simply insurance against relying on one supplier.

For operators, the immediate lesson is to keep watching inference cost and speed, not just which model scores highest on a benchmark. This is the kind of thing we build into an AI command centre, tracking where the practical economics of AI are shifting for the business.

Working With Claude field guide cover

Free Resource

Put what you just read to work

The free 32-page Working With Claude guide: the full ecosystem, Claude Code, and how to roll it out across a business.

No spam. Unsubscribe any time.

Free daily email

Get this every morning.

This brief is one item from today's AI Pulse, the short daily read we run for ourselves on what is actually happening in AI. Subscribe free and it lands in your inbox each morning.

Free daily email

Subscribe to the daily AI Pulse

One short read every morning on what is actually happening in AI. Free.

One email a day. Unsubscribe any time.