OpenAI的Jalapeño芯片在性能上超越了商用系统,AI在开发过程中发挥了关键作用,值得关注其未来的部署和应用。
OpenAI发布其首款定制推理芯片Jalapeño的性能数据,在InferenceX基准测试中,Jalapeño在每千瓦峰值吞吐量和token延迟方面均优于商用系统。AI在芯片开发中发挥了直接作用,帮助团队在九个月内完成从设计到tapeout,并优化了算术电路。Jalapeño预计将在年底前部署到OpenAI的计算基础设施中。
inference numbers published for jalapeno, team did an amazing job https://t.co/WNJIA2dESm
inference numbers published for jalapeno, team did an amazing job openai.com/index/jalapeno… tae kim @firstadopter OpenAI: "Today, we shared the first measured performance results from Jalapeño, OpenAI’s first custom inference chip. On InferenceX, a public benchmark using GPT‑OSS 120B, Jalapeño delivered more peak throughput per kilowatt and lower token latency than the commercial systems in the comparison. It also performed strongly on DeepSeek R1 and Kimi K2, showing that its gains extend across model families." "This is Jevons paradox: greater efficiency makes more uses worthwhile, expanding consumption and creating new economic activity through more work completed, better decisions, more products launched, and more revenue generated." "AI played a direct role in Jalapeño’s development, enabling the team to move from initial design to tapeout in nine months by exploring implementations, shortening design, measurement, and verification loops, and continuously iterating on model workloads. AI also helped optimize the chip’s arithmetic circuits, allowing the team to fit more compute performance into the chip on schedule." "We plan to begin deploying Jalapeño within OpenAI’s compute infrastructure by the end of the year. It is the first generation of a multigenerational roadmap: Gen 2 is deep in development, and Gen 3 is taking shape. Each generation will build on what we learn and further advance both efficiency and speed." "As we prepare for deployment, we are continuing production qualification, maturing the software, preparing to operate Jalapeño at scale, and validating performance across more models." 🔗 View Quoted Tweet 💬 3 🔄 3 ❤️ 15 👀 2216 📊 3 ⚡