ComputeLabs Research

OpenAI plans to deploy its Jalapeño custom accelerator by year-end after reporting better response speed and AI work-per-watt than GB300 in tests.

· ComputeLabs Research · from the August 25, 2026 edition

OpenAI said it plans to begin deploying Jalapeño in its compute infrastructure by the end of 2026. The custom accelerator was developed with Broadcom and is intended to support OpenAI’s artificial-intelligence models.

OpenAI chip lead Richard Ho said Jalapeño led NVIDIA’s GB300 comparison system in two tested areas: AI workload processed per unit of power and response speed. OpenAI described GB300 as the leading publicly benchmarked system used for the comparison; the supplied sources do not provide the underlying test configuration or numerical benchmark results.

The company said Jalapeño combines higher throughput with lower latency without sacrificing energy efficiency. OpenAI linked those characteristics to faster ChatGPT responses, more responsive Codex sessions and agents, and service availability as demand grows.

OpenAI will determine which of its models run on the accelerator. The company said customers could benefit from choices oriented toward either lower cost or higher performance, but it did not disclose deployment volume, manufacturing quantity or infrastructure capacity.

  • OpenAI
  • GB300

All 20 stories from August 25, 2026