OpenAI's Jalapeño chip posts first measured numbers — 1.5–1.9× more AI work per watt and 1.7–3.6× lower end-to-end latency
OpenAI published first-party benchmark numbers for Jalapeño: 1.5–1.9× more AI work per watt, 1.7–3.6× lower end-to-end latency. Vendor-tested, not independent.
OpenAI and Broadcom unveil Jalapeño, OpenAI's first custom LLM inference chip
On 2026-06-24 OpenAI and Broadcom unveiled Jalapeño, OpenAI's first custom LLM inference chip. Lab samples run today; gigawatt-scale deployment with Microsoft is planned for 2026.
NVIDIA ENPIRE: real-robot coding agents hit 99% pass@8
NVIDIA GEAR, CMU, and UC Berkeley published ENPIRE, a four-module harness that puts coding agents in a closed loop on real robots. Three frontier agents hit 99% pass@8 on five manipulation tasks.