New Delhi, Aug. 25 -- OpenAI has shared new performance results from Jalapeño, its first custom inference chip, saying the processor can deliver more AI work per watt while also reducing response times.

The chip is designed specifically for AI inference - the stage where trained models generate answers, rather than the training process used to create those models.

OpenAI says Jalapeño is designed as part of a broader system combining chips, memory, networking and software, rather than treating the processor as a standalone piece of hardware.

The company plans to begin deploying Jalapeño in its computing infrastructure by the end of 2026, with second- and third-generation chips already in development.

Jalapeño is a ...