OpenAI’s first Jalapeño benchmarks show up to 1.9x higher performance per watt and up to 3.6x lower latency across three AI models.
OpenAI and Broadcom have launched Jalapeño, their first custom AI inference processor built to improve efficiency and reduce reliance on Nvidia chips.