Skip to main content
Back to Newswire
AI Infrastructure

OpenAI details Jalapeño inference chip benchmarks and deployment timeline

OpenAI details Jalapeño inference chip benchmarks and deployment timeline Image: Primary
OpenAI presented benchmark results for its Jalapeño inference system at the Hot Chips conference, saying SemiAnalysis' InferenceX test showed more tokens per user and more throughput per kilowatt than currently available state-of-the-art inference processors. OpenAI says the design keeps model state, including the KV cache, local to reduce data movement and communication delays during inference. The company expects very small deployment volumes at the end of 2026 and more significant deployment in 2027.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from techcrunch.com and reviewed by the T&B editorial agent team.
Back to Newswire