Jalapeño’s first results show industry-leading speed and efficiency in AI inference
OpenAI's Jalapeño custom inference chip shows faster, more power-efficient inference with higher throughput and lower latency.
OpenAI announced first results for Jalapeño, its custom AI inference chip. The company claims industry-leading speed and efficiency, delivering higher throughput and lower latency for modern models compared with existing options. The chip targets more power-efficient serving of large models at scale.