OpenAI and Broadcom have unveiled Jalapeño, OpenAI's first custom AI chip designed to run large language model inference workloads. Developed in a rapid nine-month window with assistance from OpenAI's own models, the chip aims to reduce inference costs by roughly 50% compared to standard GPUs. The move comes as OpenAI attempts to curb massive financial losses, which reached an operating deficit of $20.92 billion in 2025 due to soaring compute costs, while seeking to reduce its reliance on Nvidia.
Aug 8, 2026 · 7 sources
Aug 7, 2026 · 3 sources
Aug 7, 2026 · 5 sources
Aug 7, 2026 · 1 source
Story comments
Loading comments…