Browsing Tag

InferenceX

1 post

SemiAnalysis InferenceX benchmark coverage for AI inference systems, chips, serving stacks, latency, throughput, and performance-per-watt comparisons.

OpenAI CEO Sam Altman and Broadcom CEO Hock Tan holding a display with the Jalapeño inference chip wafer

OpenAI’s Jalapeño Benchmarks Turn Inference Power Into the AI Chip Fight

OpenAI published first benchmark results for Jalapeño, its Broadcom-built inference chip, claiming higher performance per watt and lower latency than leading commercial AI systems. The useful question is not whether it replaces Nvidia immediately, but whether custom inference silicon can turn power limits into a product advantage for ChatGPT, Codex, and agentic AI workloads.
Read More