OpenAI developed the Jalapeño chip for fast inference at large scale, reports TechCrunch AI. According to the outlet's summary, the chip was tested on the InferenceX benchmark from Semianalysis.

According to this test, Jalapeño demonstrated more tokens per user and higher performance per kilowatt than solutions available at the time that were considered state-of-the-art.

The practical significance of the result depends on testing conditions: the available material lacks numerical metrics, a description of the comparison configuration, or confirmation from OpenAI. Therefore, the publication currently speaks to a claimed result from a single source rather than fully verified superiority.