AI Silicon

IBM Spyre Accelerator

What the IBM Spyre Accelerator is, how it reaches Power 11 systems, and why it is the single most confirmed signal about what the next Power generation will carry.

Spyre is IBM's PCIe AI accelerator card, built for running generative and traditional AI inference on-premises next to the transaction systems that generate the data. It matters more than most accelerator hardware for one reason: it is the subject of the clearest on-record statement IBM has made about the generation after Power 11.

Form factorPCIe add-in accelerator card
PlatformsIBM z17 mainframe and Power 11 servers
Power 11 compatibilityConfirmed for December 2025
WorkloadsOn-node generative and traditional AI inference
Related siliconTelum II, the IBM AIU chip family
Roadmap statusIBM-confirmed to continue past Power 11

The statement that makes this a roadmap signal

IBM distinguished engineer Bill Starke, the Power processor chief architect, has said on the record that Spyre "will be a piece of the Power 11 portfolio and beyond." In a landscape where IBM has published nothing about the next generation, a named architect confirming that a specific piece of silicon carries forward is unusually concrete. It is the highest-confidence item on the entire next-generation picture.

What that statement does not do is describe the next processor. It confirms an accelerator strategy continues. Everything about core counts, clocks, cache and process node for the generation after Power 11 remains unpublished.

How Spyre relates to the rest of IBM's AI silicon

Telum II ... on-chip, mainframe

IBM's mainframe processor with a native on-chip AI accelerator delivering 24 trillion operations per second, four times the AI throughput of the original Telum. Matrix multiplication and AI primitives run as native instructions rather than memory-mapped I/O. It is the z17 platform processor, and it connects to Power 11 systems by way of Spyre.

Spyre ... on-card, both platforms

Where Telum II puts inference on the processor die for the mainframe, Spyre puts it on a PCIe card that both z17 and Power 11 can take. That is what makes it the portable part of IBM's AI strategy and the part most likely to survive a processor generation change.

The AIU family ... the research lineage

IBM's Artificial Intelligence Unit chip family is the research line these accelerators descend from. Understanding it explains why IBM builds inference silicon rather than buying it, and why the design targets lower-precision formats.

Low precision by design

Telum II uses int4 and int8 formats for faster, lower-energy inference. That is a deliberate choice for enterprise scoring workloads, where a credit decision inside a transaction does not need training-grade precision and does need to finish inside the transaction.

What this means for a Power 11 purchase now

If AI inference is anywhere on your roadmap, the practical consequence is a configuration decision rather than a waiting decision. Spyre is a PCIe card. A Power 11 system specified with adapter slots left free can take one later. A system configured to the last slot cannot, without displacing something else.

Leave PCIe headroom

This is the cheapest possible hedge against an AI requirement that has not been fully scoped yet. Free adapter positions cost nothing today and are expensive to create later.

Do not wait for the next generation

Spyre works on Power 11 now. Deferring a refresh to get accelerator support on unannounced hardware gives up years of use for a capability that already exists on the current line.

Check the chassis, not just the model

Adapter capacity varies sharply across Power 11 machine types. The compact entry system has four non-hot-plug slots; the enterprise machines have four per node plus expansion drawers. The machine type decides the ceiling.

Inference next to the data is the point

The argument for on-premises accelerators over a cloud API is latency and data residency for scoring inside a live transaction. If that is not your use case, the case for the card is weaker.

Related

The full AI silicon picture is on the AI trajectory: Telum II, Spyre and AIU, and the directory entry is at Spyre Accelerator. For what this signals about the next processor, see the Power 12 CPU page and the roadmap outlook.