OpenAI 公布首款自研推理芯片 Jalapeño 的首批测试
OpenAI publishes first results for its custom Jalapeño inference chip
OpenAI 公布其首款自研推理芯片 Jalapeño 的首批测量结果,并计划于年底前开始部署。公司称,在 InferenceX 的公开模型测试中,该系统同时改善了吞吐、功耗与延迟;1.5–1.9 倍每瓦工作量及 1.7–3.6 倍更低端到端延迟等数字来自 OpenAI 的测试,应视为厂商报告,尚待更广泛的独立复现。
OpenAI has published the first measured results for Jalapeño, its first custom inference chip, and plans to begin deploying it by year-end. On public-model tests using InferenceX, the company reports gains across throughput, power, and latency. Figures such as 1.5–1.9× more work per watt and 1.7–3.6× lower end-to-end latency are OpenAI-reported and await broader independent replication.