"Read the press release for the 64 GB laptop. Every sentence about memory is about frames, not tokens."
October 2021, in Apple's Words
The M1 Pro and M1 Max were the first Apple silicon chips whose memory and bandwidth numbers would later matter to inference: 400 GB/s and 64 GB on a laptop, in 2021, when the biggest consumer graphics card had 24. Here is what Apple said the numbers were for, quoted from the release of 2021-10-18:
- "M1 Max delivers up to 400GB/s of memory bandwidth — 2x that of M1 Pro and nearly 6x that of M1 — and support for up to 64GB of unified memory."
- "M1 Pro and M1 Max also feature enhanced media engines with dedicated ProRes accelerators specifically for pro video processing."
- "M1 Max transforms graphics-intensive workflows, including up to 13x faster complex timeline rendering in Final Cut Pro compared to the previous-generation 13-inch MacBook Pro."
- "…enabling creatives like 3D artists and game developers to do more on the go than ever before."
- "…allowing playback of multiple streams of high-quality 4K and 8K ProRes video while using very little power."
And the release's only machine-learning sentence: "A 16-core Neural Engine for on-device machine learning acceleration and improved camera performance." The memory was for editors. The bandwidth was for 8K streams and 3D scenes. The Neural Engine was for the camera. Thirteen months later ChatGPT would give those same two numbers a meaning nobody in the release had in mind.
Why the Video Pitch Produced the Right Shape
This is the part of the mouse story that is not luck in the sense of coincidence; it is luck in the sense of a right answer arrived at for a different question. Professional video is also a bandwidth-and-capacity workload: an 8K ProRes stream is gigabytes per second, a colour-graded timeline with effects holds many frames resident, and a 3D scene wants its textures in memory. The engineering target — feed a wide GPU from a big, fast, shared pool without copies — is the same target a language model's decode phase would set two years later. Apple built the shape for frames, and frames and tokens turned out to want the same shape.
The clearest sign that the pitch was sincere is the media engine. ProRes accelerators — fixed-function silicon for one video codec — are an expensive commitment to a specific customer, and no one builds them as a side effect of an AI plan. The CPU track showed that this household's Ultra carries four of them. They are the fossil record of what the big-memory Max was for.
What Changes When You Read It This Way
Two things. First, the M1 Max's bandwidth is not "an AI feature" and never was: it is a video feature that AI inherited, which is why Apple's 2021 marketing has no inference claims to be held to and why this quest's vendor-claim callouts start in 2025. Second, the "admit what must be admitted" clause has a precise object: not that Apple foresaw the workload, but that Apple built the right shape for a different workload and then, once the mouse was visible, spent real silicon following it. The next three lessons are that follow-through, and what it did not fix.