Capacity · GB
What fits. A model plus its KV cache has to live in memory. 24GB runs a tuned 27B, 128GB unified runs much larger, 512GB fits frontier-class.
Hardware comparator
Every realistic way to run AI models locally, on three axes that actually decide the build: memory bandwidth, memory capacity, and what it costs you today, new and used. GPUs, Apple Silicon, unified-memory mini-PCs, and datacenter accelerators in one chart.
Prices updated 2026-09-29Indicative US market figures61 devices
The mental model
Borrowed from Ahmad Osman: local AI is capacity × bandwidth × software stack. Capacity tells you what fits, bandwidth tells you how hard the box can breathe during decode, and the stack decides how much of the spec sheet you actually cash out.
What fits. A model plus its KV cache has to live in memory. 24GB runs a tuned 27B, 128GB unified runs much larger, 512GB fits frontier-class.
How fast it generates. Decode is memory-bound, so tokens per second tracks bandwidth. GPUs stay the bandwidth kings; unified memory trades speed for size.
What you can cash out. The same silicon at stock vs tuned can differ several-fold. Fitting is not serving. The chart shows the first two; the stack is on you.
Chart
Two questions decide a local-AI box: does the model fit, and how much speed do you get for the money. So memory runs up the side (higher = more fits) and bandwidth per dollar runs across (right = more speed for your money), with bubble size showing raw bandwidth. Up and to the right is the goal. Cheap gaming cards cluster bottom-right, great value but they run out of memory; big-memory Macs and datacenter cards sit higher yet drift left, more room but slow per dollar or priced for a workstation rack. Lucebox (amber) is the box you plug in at home that holds the upper right: up to 192 GB of unified headroom with real GPU bandwidth, tuned, for a fixed price. Tip: toggle Datacenter off to see only what you can actually buy. Click Lucebox for the build.
up = more memory (what fits) · right = more bandwidth per dollar (speed for the money) · bubble = raw bandwidth · both axes log · tap Lucebox for the build
Every device
Sort any column, or tick rows to compare a few head to head. Value is capability per dollar (memory × bandwidth ÷ price), so a box with lots of fast memory for the money scores high. Price is the one that matters: new while a part is still sold, used once it is discontinued. Apple, mini-PC and Lucebox memory is the full unified pool.
| Device ⇅ | Class ⇅ | Memory GB ⇅ | Bandwidth GB/s ⇅ | Price ⇅ | Value ⇅ | |
|---|---|---|---|---|---|---|
| ★ Luceboxdetails → Lucebox Zero 495: Radeon AI PRO R9700 32GB + up to 192GB unified, tuned Lucebox engine | Lucebox | 160 u | 640 | $5,999 new | 17 | |
| MI400 / MI455X 432GB HBM4, first Helios rack shipments late Q3 2026; per-GPU price estimated | Datacenter | 432 | 19,600 | $50,000 new | 169 | |
| B200 192GB | Datacenter | 192 | 8,000 | $48,000 new | 32 | |
| B300 288GB Blackwell Ultra | Datacenter | 288 | 8,000 | $55,000 new | 42 | |
| MI325X 256GB | Datacenter | 256 | 6,000 | $26,000 new | 59 | |
| MI300X 192GB | Datacenter | 192 | 5,300 | $16,000 new | 64 | |
| H200 141GB | Datacenter | 141 | 4,800 | $37,000 new | 18 | |
| H100 80GB SXM | Datacenter | 80 | 3,350 | $32,000 new | 8 | |
| A100 80GB | Datacenter | 80 | 2,039 | $6,500 used | 25 | |
| RTX 5090 32GB Out of stock at US retailers; marketplace price vs $1,999 MSRP (GDDR7 shortage) | Consumer GPU | 32 | 1,792 | $6,450 new | 9 | |
| RTX PRO 6000 Blackwell 96GB 96GB workstation, 1792 GB/s | Workstation | 96 | 1,792 | $15,499 new | 11 | |
| RTX PRO 5000 Blackwell 72GB | Workstation | 72 | 1,344 | $4,900 new | 20 | |
| Mac Studio M5 Ultra 96GB Shipping since September 2026, 1.2 TB/s | Apple | 96 u | 1,200 | $5,499 new | 21 | |
| Mac Studio M5 Ultra 256GB 512GB ships late October | Apple | 256 u | 1,200 | $11,999 new | 26 | |
| AMD Instinct MI50 32GB Cheap-VRAM favorite, used only | Datacenter | 32 | 1,024 | $320 used | 102 | |
| RTX 3090 Ti 24GB | Consumer GPU | 24 | 1,008 | $1,300 used | 19 | |
| RTX 4090 24GB | Consumer GPU | 24 | 1,008 | $2,500 used | 10 | |
| RTX 5080 16GB | Consumer GPU | 16 | 960 | $1,650 new | 9 | |
| RX 7900 XTX 24GB | Consumer GPU | 24 | 960 | $900 used | 26 | |
| RTX 6000 Ada 48GB | Workstation | 48 | 960 | $6,800 new | 7 | |
| RTX 3090 24GB 24GB used-market favorite for local AI | Consumer GPU | 24 | 936 | $1,350 used | 17 | |
| RTX 3080 12GB | Consumer GPU | 12 | 912 | $340 used | 32 | |
| Intel Arc Pro B60 Dual 48GB Two B60 GPUs on one card | Workstation | 48 | 912 | $1,200 new | 36 | |
| Tesla V100 32GB | Datacenter | 32 | 900 | $700 used | 41 | |
| RTX 5070 Ti 16GB | Consumer GPU | 16 | 896 | $1,200 new | 12 | |
| Radeon PRO W7900 48GB | Workstation | 48 | 864 | $3,897 new | 11 | |
| Radeon PRO W7800 48GB Gigabyte AI TOP 48GB variant | Workstation | 48 | 864 | $2,700 new | 15 | |
| L40S 48GB | Datacenter | 48 | 864 | $8,500 new | 5 | |
| Mac Studio M3 Ultra 96GB Replaced by the M5 Ultra Mac Studio | Apple | 96 u | 819 | $3,300 used | 24 | |
| Mac Studio M3 Ultra 512GB 512GB retired; the M5 Ultra 512GB ships late October | Apple | 512 u | 819 | $30,000 used | 14 | |
| RX 7900 XT 20GB | Consumer GPU | 20 | 800 | $700 used | 23 | |
| Mac Studio M2 Ultra 192GB | Apple | 192 u | 800 | $5,500 used | 28 | |
| RTX A6000 48GB | Workstation | 48 | 768 | $3,500 used | 11 | |
| RTX 3080 10GB | Consumer GPU | 10 | 760 | $370 used | 21 | |
| RTX 4080 Super 16GB | Consumer GPU | 16 | 736 | $1,100 used | 11 | |
| RTX 4080 16GB | Consumer GPU | 16 | 717 | $1,080 used | 11 | |
| RTX 4070 Ti Super 16GB | Consumer GPU | 16 | 672 | $770 used | 14 | |
| RTX 5070 12GB | Consumer GPU | 12 | 672 | $1,000 new | 8 | |
| RX 9070 XT 16GB | Consumer GPU | 16 | 640 | $800 new | 13 | |
| RX 9070 16GB | Consumer GPU | 16 | 640 | $690 new | 15 | |
| Radeon AI PRO R9700 32GB | Workstation | 32 | 640 | $1,800 new | 11 | |
| RX 7800 XT 16GB | Consumer GPU | 16 | 624 | $500 used | 20 | |
| RTX 2080 Ti 22GB (modded) China-modded 22GB refurb | Consumer GPU | 22 | 616 | $550 used | 25 | |
| MacBook Pro M5 Max 128GB | Apple | 128 u | 614 | $6,749 new | 12 | |
| Tenstorrent Wormhole n300 24GB RISC-V Tensix, fully OSS stack | Workstation | 24 | 576 | $1,399 new | 10 | |
| Mac Studio M4 Max 128GB 40-core; 128GB pulled, max now 96GB | Apple | 128 u | 546 | $3,200 used | 22 | |
| RX 6800 16GB | Consumer GPU | 16 | 512 | $380 used | 22 | |
| Tenstorrent Blackhole p150 32GB RISC-V Blackhole, fully OSS stack | Workstation | 32 | 512 | $1,399 new | 12 | |
| RTX 4070 12GB | Consumer GPU | 12 | 504 | $500 used | 12 | |
| Intel Arc Pro B60 24GB | Workstation | 24 | 456 | $674 new | 16 | |
| RTX 3060 Ti 8GB | Consumer GPU | 8 | 448 | $240 used | 15 | |
| Mac Studio M4 Max 36GB 32-core GPU, ~410 GB/s; replaced by the M5 Max Mac Studio | Apple | 36 u | 410 | $2,100 used | 7 | |
| RTX 3060 12GB | Consumer GPU | 12 | 360 | $300 used | 14 | |
| Tesla P40 24GB Cheap 24GB, no video out, needs blower | Datacenter | 24 | 347 | $270 used | 31 | |
| MacBook Pro M5 Pro 64GB | Apple | 64 u | 307 | $3,499 new | 6 | |
| Mac mini M4 Pro 64GB 64GB config pulled (DRAM shortage) | Apple | 64 u | 273 | $2,200 used | 8 | |
| DGX Spark 128GB (GB10) Coherent memory + CUDA | Mini-PC | 128 u | 273 | $6,500 new | 5 | |
| Ryzen AI Max+ 395 128GB Strix Halo, ~96GB allocatable to GPU | Mini-PC | 128 u | 256 | $3,499 new | 9 | |
| Jetson AGX Orin 64GB 275 TOPS edge dev kit; was $1,999 before July 2026 | Mini-PC | 64 u | 204 | $3,499 new | 4 | |
| MacBook Air M5 32GB | Apple | 32 u | 153 | $1,499 new | 3 | |
| Mac mini M4 16GB Replaced by the M6 Mac mini in September 2026 | Apple | 16 u | 120 | $850 used | 2 |
Method. Memory bandwidth is peak theoretical. Each device shows the price that actually applies: the current new price while it is still sold, the current used price once it is discontinued. Figures are indicative US-market numbers triangulated across retailers (Newegg, Amazon, B&H, Apple) and used marketplaces (eBay sold listings, resellers) as of 2026-09-29; the 2026 memory shortage has pushed new prices up and pulled several high-memory configs from sale. Value = memory × bandwidth ÷ price (capability per dollar). Datacenter parts are quote-based and sold in 8-GPU boards, so their per-unit figures are approximate. "u" marks unified memory. Numbers move fast; treat them as a snapshot, not a quote.
Our build
A Radeon AI PRO R9700 with 32GB GDDR6 and 640 GB/s bandwidth, paired with 128GB unified memory and a tuned stack. Plug-and-play, validated, and warrantied. The numbers above are here so you can judge the configuration directly.
Next to the others
The comparison page puts Lucebox next to cloud APIs, a DGX Spark and a Mac Studio, row by row.