OpenAI just revealed the first results from Jalapeño, its first custom AI inference chip, and the gains are huge.
Across big models including DeepSeek R1 670B and Kimi K2.5 1T, it delivers 1.5-1.9x more AI work per watt, 1.7–3.6x lower end to end latency, and up to 4.1x higher performance for highly interactive workloads versus comparison systems.
AI helped design the chip, enabling OpenAI to go from initial design to tapeout in only 9 months.
AI generated implementations for selected workloads were already 1.5–1.8x faster than human expert written versions.
AI designs better chips → better chips run AI faster and cheaper → stronger AI designs even better chips → repeat
顯示更多