← All news·2026-09-19·3 min read

OpenAI Unveils Jalapeño: LLM-Designed AI Chip With 3.6x Lower Latency Than Nvidia GB300

OpenAI has unveiled Jalapeño, an AI accelerator with inference latency up to 3.6 times lower than the Nvidia GB300. A team of fewer than 100 engineers built it in 20 months using OpenAI's own language models.

aihardwareopenaichip-design

OpenAI has unveiled Jalapeño, its own AI accelerator. A team of fewer than 100 engineers built it in 20 months. Inference latency is up to 3.6 times lower than the Nvidia GB300. OpenAI's own language models were among the tools used, helping to optimize the circuit designs.

The most telling detail: LLMs raised the performance of a test compute kernel from 0.31% to 88.94% of the theoretical maximum in about 40 hours. This is not "chips designing themselves." It is engineers with a tool that handles the most labor-intensive part of parameter tuning.

Source: spectrum.ieee.org

Free course

Stop reading about AI — start building with it

The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.

Start free →
EAEvgenii Arsentev

Author

Evgenii Arsentev

PhD · Chief Executive Officer, digital health