0
Applied AI·August 25, 2026·1 min read

OpenAI’s Jalapeno chip outperformed the GB300 on power and speed, according to OpenAI

Share

If Jalapeno really delivers better work-per-watt and latency than GB300 for inference, the cost curve for serving large models just bent again. For operators, that means re-running your unit economics assumptions on where to host inference and how aggressively to lean into heavier models at the edge of your UX.