OpenAI built a chip in nine months. Then it let AI rewrite the code.
The New Stack Amanda Caswell
OpenAI published its first performance results for Jalapeño, a custom inference chip built with Broadcom and tested across multiple large language models. Jalapeño delivered 1.5 to 1.9 times more work per watt and cut end-to-end latency by 1.7 to 3.6 times, while the chip is rated at 700 watts and did not exceed 550 watts in these tests. OpenAI says it will start using Jalapeño in its own infrastructure by the end of the year and is already working on the next two generations.
Why it matters
When OpenAI unveiled Jalapeño, its first custom inference chip, in June, the company made some big promises. The chip, developed The post OpenAI built a chip in nine months. Then it let AI rewrite the code. appeared first on The New Stack.