OpenAI Publishes Benchmark Results for Custom Inference Chip Jalapeno
OpenAI held a briefing on Jalapeño, a custom inference chip co-developed with Broadcom, and published benchmark results. The chip reportedly achieves both high throughput and low latency in a single architecture, a combination that typically requires a trade-off. OpenAI plans to deploy the first generation within its internal computing infrastructure by the end of the year.