OpenAI表示,其Jalapeño芯片能够比竞争对手提供更快的AI响应速度。

qimuai 发布于 阅读:65 一手编译

OpenAI表示,其Jalapeño芯片能够比竞争对手提供更快的AI响应速度。

内容来源:https://www.theverge.com/ai-artificial-intelligence/984290/openai-jalapeno-ai-chip-benchmarks

内容总结:

OpenAI发布自研AI芯片“Jalapeño”,称性能全面超越英伟达旗舰产品

OpenAI于本周二发布博文宣布,其自主研发的新型AI芯片“Jalapeño”在任务处理效率和响应速度上均优于其他AI系统。OpenAI硬件副总裁理查德·何(Richard Ho)在记者会上表示,该芯片实现了“两全其美”——既降低了延迟,又提升了吞吐量,而传统AI系统通常不得不在二者之间做出权衡。

据悉,Jalapeño芯片于今年6月首次亮相,是与博通合作生产的专用集成电路(ASIC),专为AI推理设计,即运行已训练好的模型以完成任务或部署智能体的过程。在基于InferenceX基准平台的测试中,Jalapeño的表现超越了当时成绩最佳的英伟达GB200或GB300超级芯片。测试数据显示,在对GPT-OSS 120B、DeepSeek R1及Kimi K2.5 1T三款模型的推理任务中,Jalapeño每瓦特产生的AI工作量是参照系统的1.5至1.9倍,端到端延迟则降低1.7至3.6倍。何表示,这意味着用户将获得“更快的响应、更灵敏的智能体以及高需求下更稳定的服务”。

OpenAI计划于今年底小规模部署Jalapeño芯片,并预计在2027年大幅提升产量,但未透露明年的具体部署数量。何同时强调,公司不会用Jalapeño完全替代现有芯片产品线,整体算力战略仍将包括英伟达等“优质合作伙伴”。此外,OpenAI已着手推进该芯片的第二代和第三代研发工作。

中文翻译:

OpenAI表示,其新款AI芯片Jalapeño在执行任务时效率更高,响应速度也快于其他AI系统,这是根据周二发布的一篇博文得出的结论。在与记者的简报会上,OpenAI硬件副总裁理查德·何表示,Jalapeño兼具“两全其美”的优势——延迟更低、吞吐量更高,而AI系统通常“不得不在两者之间做出取舍”。

OpenAI称其Jalapeño芯片能以比竞争对手更快的速度驱动AI响应

Jalapeño在AI推理基准测试中表现优于英伟达的超强芯片。

Jalapeño在AI推理基准测试中表现优于英伟达的超强芯片。

Jalapeño于今年6月首次亮相,是一款与博通合作打造的专用集成电路(ASIC)。它专为AI推理而设计——即运行训练好的AI模型以完成任务或部署智能体的过程。

为了衡量Jalapeño的性能,OpenAI使用了InferenceX这一基准测试平台,该平台能展示AI系统处理推理任务的能力。测试将Jalapeño的表现与当时记录的最佳结果进行了对比,后者来自英伟达的GB200或GB300超强芯片。OpenAI表示,在GPT-OSS 120B、DeepSeek R1和Kimi K2.5 1T三个模型上,Jalapeño每瓦特完成的AI工作量是对照系统的1.5至1.9倍,同时在这三个模型上的端到端延迟比对照系统低1.7至3.6倍。这意味着该芯片能为用户提供“更快的响应、更灵敏的智能体,以及随着需求增长更可靠的访问体验”,何表示。

OpenAI计划在今年年底前以“小批量”部署Jalapeño,但将在2027年开始“逐步扩大产量”,何补充道。不过,该公司并未透露明年的计划部署芯片数量。

尽管有这些性能提升,何表示OpenAI并不打算用Jalapeño全面替换其现有芯片产品线,称其整体计算策略包括“非常优秀的合作伙伴”,如英伟达。OpenAI将继续开发这款新芯片的第二代和第三代产品。

热门推荐

英文来源:

OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday. During a briefing with reporters, OpenAI hardware vice president Richard Ho said Jalapeño offers the “best of both worlds” with lower latency and higher throughput, as AI systems typically “have to make a trade-off between the two.”
OpenAI says its Jalapeño chip can power faster AI responses than the competition
Jalapeño outperformed Nvidia’s superchips on an AI inference benchmark test.
Jalapeño outperformed Nvidia’s superchips on an AI inference benchmark test.
First introduced in June, Jalapeño is an Application-Specific Integrated Circuit (ASIC) made in partnership with Broadcom. It’s designed for AI inference — the process of running a trained AI model to complete a task or deploy an agent.
To measure Jalapeño’s performance, OpenAI used InferenceX, a benchmarking platform that shows how well AI systems handle inference. The test compared Jalapeño’s performance against the best results recorded at the time, which were with Nvidia’s GB200 or GB300 superchips. OpenAI says Jalapeño delivered 1.5 to 1.9 times more AI work per watt across GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T than the comparison systems, while offering 1.7 to 3.6 times lower end-to-end latency across the three models. That means the chip can provide users with “faster responses, more responsive agents, and more reliable access as the demand grows,” according to Ho.
OpenAI plans to deploy Jalapeño in “small volumes” by the end of this year, but will begin to “ramp the volume up” into 2027, Ho added. The company doesn’t say how many chips it plans to deploy next year, however.
Even with these performance improvements, Ho said OpenAI doesn’t expect to replace its entire chip lineup with Jalapeño, saying its overall compute strategy includes “very good partners,” like Nvidia. OpenAI will continue developing the second and third generations of the new chip.
Most Popular

ThevergeAI大爆炸

文章目录


    扫描二维码,在手机上阅读