
New hardware shows efficiency gains over Nvidia Blackwell in early testing, with full deployment slated for 2027.
AI-generated summary
Jalapeño was first announced in October and is being developed by OpenAI in collaboration with Broadcom.
At the Hot Chips conference on Tuesday, OpenAI shared a more detailed look at Jalapeño, including the first batch of benchmark results for the new system. Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the-art inference processors.
“The bottom line is that the results show a very, very significant performance advance over state of the art,” said Richard Ho, OpenAI’s head of hardware, in a press call. “Jalapeño can serve more AI work per unit of power, while also returning responses more quickly. It’s very efficient to serve a lot of customers, but it can also be very low latency.”
Notably, that comparison is against an Nvidia Blackwell system — but by the time Jalapeño reaches full deployment, the competition may have advanced significantly. Ho estimated that Jalapeño would deploy at the end of 2026 “in very small volumes,” with more significant deployment coming in 2027.
First announced last October, Jalapeño was developed by OpenAI in close collaboration with Broadcom, with OpenAI’s own models assisting in the development process. The company plans to make Jalapeño a multigenerational platform, allowing AI products, models, chips, and memory all developed in concert.
Because of that full-stack approach, OpenAI was able to address specific phases in the inference process that often cause friction during inference processing. In particular, Jalapeño is designed to minimize delays during the prefill and communication phases of processing, which OpenAI says often act as bottlenecks.
“We designed Jalapeño to minimize data movement and communication delays,” the company said in a blog post presenting the results. “This means that model state, including the KV cache used while generating a response, can be explicitly placed and kept local while the system activates the right combination of compute, memory, and networking for each inference phase.”
AI outlook — possibilities, not facts
Small volume deployment of Jalapeño chips by end of 2026.
Likely · Within months

Apple launched a new Mac Mini featuring the M6 processor, starting at $899, and an M5 Pro variant starting at $1,699. The company also updated the Mac Studio with M5 Max and M5 Ultra chips. The new devices ship September 22, featuring macOS 27 and Siri AI.

Fitbit co-founders James Park and Eric Friedman have launched Luffu Link, a $250 LTE-enabled wearable designed for family caregiving. The device tracks health metrics, enables voice logging, and provides location sharing to help families monitor well-being without constant check-ins.

Apple has refreshed its Mac mini and Mac Studio desktops, introducing the 2nm M6 chip and the high-capacity M5 Ultra. The updates emphasize support for local AI inference and distributed machine learning workflows, catering to developers and researchers.

OpenAI has introduced Jalapeño, a new AI-focused ASIC developed with Broadcom. The chip reportedly delivers higher efficiency and lower latency than Nvidia's GB200/GB300 superchips, with small-scale deployment planned for late 2024 and scaling through 2027.

Uber rolls out a live video streaming feature for parents/guardians to monitor teens' rides via the driver's selfie camera, set to launch nationwide in the U.S. within weeks.

Smart monitors, with internet connectivity and operating systems, are increasingly pushing ads and tracking users, mirroring the privacy concerns seen in smart TVs. Manufacturers like LG and Samsung are using ad-serving OSes, potentially leveraging automatic content recognition (ACR) to track activity, raising significant privacy risks.