OpenAI Unveils Its Own Jalapeño Chip: An Accelerator 1.5–2 Times More Efficient Than Nvidia

cryptonews.ruPublicado em 2026-08-30Última atualização em 2026-08-30

Resumo

On August 25, 2026, OpenAI unveiled test results for its custom inference accelerator, Jalapeño. On the public InferenceX benchmark, systems using the new chip delivered 1.5–1.9x more computations per watt at peak throughput and reduced response latency by 1.7–3.6x compared to systems based on Nvidia GB200 and GB300. The chip, rated at 700W nominal power, consumed up to 550W under tested loads, with a block of 128 units reaching 1.7 exaflops in 4-bit precision. OpenAI plans to deploy Jalapeño in its infrastructure by late 2026, marking the first generation of a multi-year platform. Jalapeño was developed in collaboration with Broadcom (for the die and networking) and Celestica (for boards and racks) over nine months. It is specifically designed for OpenAI's own predictable, high-volume inference workloads around its language models, computational kernels, and data movement, unlike Nvidia's general-purpose accelerators. The initiative is backed by an agreement with Broadcom to deploy 10 GW of OpenAI's custom accelerators between late 2026 and 2029. While OpenAI will continue relying on external suppliers like Nvidia for training cutting-edge models and parts of inference, the company aims to control the processor architecture for its largest daily computational stream. This shift addresses the economics of inference: by building chips as internal components, OpenAI avoids paying the market premium associated with Nvidia's high-margin commercial GPUs, directly lowering the co...

On August 25, 2026, OpenAI published the first test results of its own inference accelerator, Jalapeño. On the public InferenceX benchmark by SemiAnalysis, systems equipped with the new chip demonstrated 1.5–1.9 times more computations per watt at peak throughput and reduced response latency by 1.7–3.6 times compared to systems based on Nvidia GB200 and GB300. Tests were conducted on the GPT-OSS-120B, DeepSeek R1, and Kimi K2.5 models. For a company that daily serves the massive traffic of ChatGPT, Codex, and its own API, such metrics are directly tied to the cost per answer.

The chip is rated for a nominal power of 700W, while its sustained power consumption under tested loads remained at up to 550W. A block of 128 Jalapeño chips is claimed to deliver 1.7 exaflops of compute in 4-bit format.

Where Independence from Nvidia Ends

Jalapeño was co-developed with partners: Broadcom was responsible for the silicon implementation and networking aspects, while Celestica handled the boards and racks. From initial design to tape-out took nine months. Deployment of the chip in OpenAI's infrastructure is planned for the end of 2026, with Jalapeño becoming the first generation in a multi-year platform set for future expansion.

Nvidia's general-purpose accelerator must serve a broad range of clients and task types. Jalapeño, however, was engineered for the specific workload that OpenAI itself generates, measures, and can modify alongside its software stack—centered around its own language models, compute cores, data movement, memory, and network exchange.

Scale Set for 10 GW

The foundation for the entire program is an agreement between OpenAI and Broadcom, under which the parties committed to deploying 10 GW of OpenAI's own designed accelerators. Rack installations will begin in the second half of 2026 and continue through the end of 2029. At this scale, differences in watts and milliseconds cease to be lab metrics and become cost factors for electricity, cooling, memory, network, racks, and data center floor space.

At the same time, OpenAI is not abandoning external suppliers entirely:

  • Training cutting-edge models still requires external accelerators;
  • Volume production of Jalapeño depends on partners for manufacturing, memory, packaging, and assembly;
  • The company will continue to extensively use chips from Nvidia and other suppliers—both for training and for part of the inference.

The project's boundary is therefore clear: OpenAI maintains an external supply chain but takes control of the hardware architecture for its largest stream of daily compute.

The Economics of Inference

The difference between a custom chip and purchasing GPUs extends beyond energy consumption. Nvidia concluded its 2026 fiscal year with revenue of $215.9 billion and a gross margin of 71.1%, while its data center segment grew 68% year-over-year—these figures are reflected in the company's filings with the Securities and Exchange Commission (SEC). A GPU buyer pays not only for the hardware but also for the vendor's commercial profit margin on top of it.

OpenAI builds the chip as an internal component of its own service and benefits from lowering the total cost of processing a request, even without a separate market markup on the processor itself. Part of the cost that previously accrued to the supplier of general-purpose accelerators is now being turned into the company's own savings and additional compute capacity.

Cheaper and faster inference enables longer-running agentic tasks, servicing more concurrent requests, and reducing product costs without proportional growth in data centers. Increased service usage, in turn, boosts the utilization of the custom platform, making the next generation of specialized chips more economically justified.

Jalapeño changes OpenAI's position in the AI infrastructure chain: the company already controlled the model, software stack, and product, and now gains control over the processor architecture that determines the price of a response. Nvidia retains the market for model training and general-purpose accelerators, but the most massive part of OpenAI's workload is no longer a guaranteed sale for them.

AI Opinion

From the perspective of data-driven analysis, the Jalapeño case follows a trajectory already taken by other industry leaders before OpenAI. Anthropic committed to a million TPUs from Google back in October 2025, and Midjourney shifted image generation to Google Cloud TPUs for cost savings. The underlying logic is the same: hyperscalers move away from general-purpose GPUs to chips tailored for their own workload once the volume of inference becomes sufficiently predictable and massive.

However, the strategy also has a downside. Specialization for a specific model and stack reduces flexibility when AI architectures evolve, and reliance on Broadcom and Celestica for manufacturing creates a single point of risk for the entire nine-month development cycle. Will Jalapeño remain competitive after two or three generations of models, or will the narrow optimization for today's inference become a limitation tomorrow?

end-content

Criptomoedas em alta

Perguntas relacionadas

QWhat are the key performance advantages of OpenAI's Jalapeño inference accelerator compared to Nvidia GB200/GB300 systems according to the SemiAnalysis benchmark?

AAccording to the SemiAnalysis InferenceX benchmark, systems using OpenAI's Jalapeño chip delivered 1.5–1.9x more computations per watt at peak throughput and reduced response latency by 1.7–3.6x compared to systems based on Nvidia GB200 and GB300.

QWhat is the planned scale of OpenAI's custom accelerator program with Broadcom and what is the associated timeline?

AOpenAI and Broadcom have an agreement to deploy 10 GW of OpenAI's custom accelerators. Rack deployment is scheduled to begin in the second half of 2026 and last until the end of 2029.

QWhat are the main business and strategic reasons for OpenAI to develop its own inference chip, as mentioned in the article?

AThe main reasons are reducing the total cost per query for OpenAI's massive daily inference workload (like ChatGPT) and internalizing a cost center that was previously a vendor's profit margin. A cheaper, faster inference allows for more complex tasks, more concurrent requests, and lower product costs without proportional datacenter expansion, creating a positive feedback loop for future specialized chip development.

QHow does the development and production model of the Jalapeño chip differ from the typical Nvidia GPU model, according to the article?

AThe Jalapeño chip is designed for a specific, known workload generated by OpenAI's own models and software stack, allowing for deep optimization. In contrast, Nvidia's universal accelerators must serve a wide range of clients and task types. OpenAI relies on partners like Broadcom and Celestica for production, but controls the chip architecture to optimize for its own operational costs rather than selling the chip as a standalone product with a market markup.

QWhat are the potential downsides or risks associated with OpenAI's strategy of developing a highly specialized chip like Jalapeño?

AThe strategy carries risks, including reduced flexibility if AI model architectures change, as the chip is optimized for today's specific inference patterns. Furthermore, dependence on manufacturing partners like Broadcom and Celestica creates a single point of risk across the nine-month development cycle. The article questions whether Jalapeño will remain competitive in 2-3 model generations or if its narrow optimization will become a future limitation.

Leituras Relacionadas

ECB's Schnabel says central bank money 'should move to blockchain'

ECB Executive Board member Isabel Schnabel stated that central bank money "must move onto the blockchain." She argued that stablecoins lack the independent ability to scale liquidity during financial stress—a gap only a central bank can fill. Her proposed solution involves tokenization, which she says can make transactions faster, safer, and more programmable, but only if the safest asset (central bank money) is on the same "rails" as other tokenized assets. This marks a notable shift for the Eurosystem, which had previously viewed Distributed Ledger Technology (DLT) mainly as a tool for regulating stablecoins and crypto, not as infrastructure to adopt directly. The first step is Project Pontes, launching in September. It will initially synchronize the ECB’s existing TARGET services with private DLT platforms. Eventually, it aims to enable settlement finality on a Eurosystem-managed DLT platform with smart contract functionality and 24/7 operation. The long-term strategy is Project Appia, tasked with developing the architecture, standards, and legal framework for a genuine European tokenized asset market by 2028. Trials have already processed around €1.6 billion, and since March 2026, the ECB accepts DLT-based assets as collateral. While not directly impacting Bitcoin's price, these developments signal that a major G7 central bank is preparing to settle transactions on-chain, lending legitimacy to the underlying infrastructure of crypto markets. The ECB's move to avoid "disintermediation" by private tokenization shows that debates in central bank boardrooms are now aligning with discussions long followed in the crypto space.

cryptonews.ruHá 1h

ECB's Schnabel says central bank money 'should move to blockchain'

cryptonews.ruHá 1h

Ripple Labs Warns of a New Challenge for Blockchains

Ripple Labs is preparing the XRP Ledger for quantum threats but now emphasizes a broader challenge: ensuring financial infrastructure can adapt simultaneously to quantum computing and artificial intelligence. Senior Director of Engineering Ayo Akinyele states the goal is not merely anticipating a "Q-Day" but building infrastructure capable of preemptively adopting new security mechanisms without network disruption. The shift to post-quantum cryptography involves more than swapping algorithms; it requires flexible infrastructure, improved key management, and clear upgrade paths. Financial systems were not designed with quantum computers in mind, necessitating a rethink of transaction, identity, asset, and data protection. This is underscored by significant investments, such as the $2 billion U.S.-IBM quantum factory initiative and a 2030 U.S. government mandate for post-quantum cryptography adoption. AI introduces distinct risks by driving automation and autonomous economic activity, increasing demand for an always-on, internet-oriented payment infrastructure. AI agents could soon conduct transactions independently, creating new security requirements. Ripple Labs advocates for proactive, orderly preparation rather than a crisis-driven transition, having outlined a four-phase strategy aiming for a full XRP Ledger transition by 2028, now expanded to address both quantum and AI-driven challenges.

cryptonews.ruHá 1h

Ripple Labs Warns of a New Challenge for Blockchains

cryptonews.ruHá 1h

Trading

Spot

Artigos em Destaque

Como comprar CHIP

Bem-vindo à HTX.com!Tornámos a compra de USD.AI (CHIP) simples e conveniente.Segue o nosso guia passo a passo para iniciar a tua jornada no mundo das criptos.Passo 1: cria a tua conta HTXUtiliza o teu e-mail ou número de telefone para te inscreveres numa conta gratuita na HTX.Desfruta de um processo de inscrição sem complicações e desbloqueia todas as funcionalidades.Obter a minha contaPasso 2: vai para Comprar Cripto e escolhe o teu método de pagamentoCartão de crédito/débito: usa o teu visa ou mastercard para comprar USD.AI (CHIP) instantaneamente.Saldo: usa os fundos da tua conta HTX para transacionar sem problemas.Terceiros: adicionamos métodos de pagamento populares, como Google Pay e Apple Pay, para aumentar a conveniência.P2P: transaciona diretamente com outros utilizadores na HTX.Mercado de balcão (OTC): oferecemos serviços personalizados e taxas de câmbio competitivas para os traders.Passo 3: armazena teu USD.AI (CHIP)Depois de comprar o teu USD.AI (CHIP), armazena-o na tua conta HTX.Alternativamente, podes enviá-lo para outro lugar através de transferência blockchain ou usá-lo para transacionar outras criptomoedas.Passo 4: transaciona USD.AI (CHIP)Transaciona facilmente USD.AI (CHIP) no mercado à vista da HTX.Acede simplesmente à tua conta, seleciona o teu par de trading, executa as tuas transações e monitoriza em tempo real.Oferecemos uma experiência de fácil utilização tanto para principiantes como para traders experientes.

683 Visualizações TotaisPublicado em {updateTime}Atualizado em 2026.06.02

Como comprar CHIP

Discussões

Bem-vindo à Comunidade HTX. Aqui, pode manter-se informado sobre os mais recentes desenvolvimentos da plataforma e obter acesso a análises profissionais de mercado. As opiniões dos utilizadores sobre o preço de CHIP (CHIP) são apresentadas abaixo.

活动图片