"A generation number tells you when. It does not tell you how much."
The Ladder by Year
| Generation | First shipped | Process (Apple's words) | Base-chip bandwidth | Top-tier bandwidth | Evidence |
|---|---|---|---|---|---|
| M1 | 2020-11 | 5 nm | not published (~68 GB/s by our fit) | M1 Ultra 800 GB/s (2022-03) | vendor + physics |
| M2 | 2022-06 | second-generation 5 nm | 100 GB/s | M2 Ultra 800 GB/s (2023-06) | vendor |
| M3 | 2023-10 | 3 nm | 100 GB/s | M3 Ultra 819 GB/s (2025-03) | vendor |
| M4 | 2024-05 | second-generation 3 nm | 120 GB/s | M4 Max 546 GB/s (no Ultra) | vendor |
| M5 | 2025-10 | third-generation 3 nm | 153 GB/s | M5 Ultra "1.2TB/s" (announced 2026-08-25, ships 2026-09-22) | vendor |
| M6 | announced 2026-08-25 | 2 nm — "Apple's first state-of-the-art 2-nanometer chip" | not published | — | vendor (announcement only) |
Two honest readings of that table. First, the base chip's bandwidth roughly doubled across five generations (68 → 153 GB/s) while the top tier grew by about half (800 → 1,229 GB/s). The dramatic numbers in Apple silicon come from tiers, not from generations: an M1 Ultra from 2022 still out-feeds an M5 Max from 2026. Second, the memory technology drove the generation gains: LPDDR4X to LPDDR5 (M1→M2), then LPDDR5X at 7500, 8533 and 9600 MT/s (M4, M4 Pro/Max, M5). The process node shrank once in the same period (5 nm to 3 nm at M3) and then went through two revisions of 3 nm, and none of it moved the bandwidth number, because bandwidth lives in the memory and the bus, not in the transistors.
What Changed Shape in the M5 Generation
Two things, both vendor-described and neither measured in this quest. The M5 Pro and M5 Max are built from two third-generation 3 nm dies — Apple calls it the Fusion Architecture and says it "connects two dies into a single SoC" using advanced packaging; the per-die contents are not disclosed. The M5 Ultra then joins two of those dual-die Max chips with UltraFusion into "the quad-die architecture — a first for Apple silicon", with inter-die bandwidth "over 4.4TB/s". So the die count per tier went 1/1/1/2 (M3 family) to 1/2/2/4 (M5 family). Keep that in mind for the next lesson, where "monolithic" gets three meanings.
The M5 generation also introduced per-core Neural Accelerators in the GPU and a new core naming — "super cores" plus "performance cores", with no efficiency cores listed for Pro, Max and Ultra. The CPU track covers what that means; here it is enough to note that Apple's own M5 numbers attribute prefill gains to the accelerators and decode gains to bandwidth, which is exactly the split this quest is built around.
Vendor Claims, With Their Baselines
| Claim | Stage it measures | Baseline | Source |
|---|---|---|---|
| M5: "over 4x the peak GPU compute performance compared to M4" | peak metric, not a workload | M4 | Apple newsroom, 2025-10-15 |
| M5 Pro/Max: "up to 4x faster LLM prompt processing than M4 Pro and M4 Max" | prefill | M4 Pro / M4 Max | Apple newsroom, 2026-03-03 |
| M5 Ultra: "up to 4.3x the peak AI compute performance when compared to M3 Ultra" | peak metric | M3 Ultra 32C/80G 512 GB, tested July 2026 | Mac Studio release, 2026-08-25 |
| M5 Ultra: "up to 4.5x the peak GPU compute for AI compared to M3 Ultra" | peak metric | M3 Ultra 32C/80G 256 GB, tested August 2026 | M6/M5 Ultra release, same day |
| M5 Ultra: "up to 4x faster [LLM prompt processing] than M3 Ultra" in LM Studio | prefill | M3 Ultra | Mac Studio release, 2026-08-25 |
Notice the two same-day releases disagree — 4.3x against 4.5x for what reads as the same comparison, with different baseline configurations and test months in the footnotes. That is not a scandal; it is what "peak" means. It is also why this quest never copies a vendor multiplier without the stage and the baseline beside it, and why the M5 generation appears here as claims only: the household's newest Mac Studio is an M3 Ultra, and no M5 number in this quest is measured. When an M5 Ultra is measured, that is an evolve, not a guess.