Route the Work, Not Just the Data: GPUs, CPUs, and the Rise of AI-Native SSDs
Paul Woll · August 18, 2026
An LLM request is many kinds of work, and only some of it needs a GPU. Why the winning AI architecture routes each operation to the cheapest tier that can perform it, from today's SSD-backed KV-cache tiers to a proposed five-plane AI-native storage device, with a laptop-reproducible 102 GB data-movement benchmark and a falsifiable proposal for computational-storage hardware.
Research repo