Hardware comparator

Local AI hardware, compared.

Every realistic way to run AI models locally, on three axes that actually decide the build: memory bandwidth, memory capacity, and what it costs you today, new and used. GPUs, Apple Silicon, unified-memory mini-PCs, and datacenter accelerators in one chart.

Prices updated 2026-09-29Indicative US market figures61 devices

The mental model

Which bottleneck am I buying?

Borrowed from Ahmad Osman: local AI is capacity × bandwidth × software stack. Capacity tells you what fits, bandwidth tells you how hard the box can breathe during decode, and the stack decides how much of the spec sheet you actually cash out.

01

Capacity · GB

What fits. A model plus its KV cache has to live in memory. 24GB runs a tuned 27B, 128GB unified runs much larger, 512GB fits frontier-class.

02

Bandwidth · GB/s

How fast it generates. Decode is memory-bound, so tokens per second tracks bandwidth. GPUs stay the bandwidth kings; unified memory trades speed for size.

03

Software stack

What you can cash out. The same silicon at stock vs tuned can differ several-fold. Fitting is not serving. The chart shows the first two; the stack is on you.

Chart

How to read this chart.

Two questions decide a local-AI box: does the model fit, and how much speed do you get for the money. So memory runs up the side (higher = more fits) and bandwidth per dollar runs across (right = more speed for your money), with bubble size showing raw bandwidth. Up and to the right is the goal. Cheap gaming cards cluster bottom-right, great value but they run out of memory; big-memory Macs and datacenter cards sit higher yet drift left, more room but slow per dollar or priced for a workstation rack. Lucebox (amber) is the box you plug in at home that holds the upper right: up to 192 GB of unified headroom with real GPU bandwidth, tuned, for a fixed price. Tip: toggle Datacenter off to see only what you can actually buy. Click Lucebox for the build.

Show

up = more memory (what fits)  ·  right = more bandwidth per dollar (speed for the money)  ·  bubble = raw bandwidth  ·  both axes log  ·  tap Lucebox for the build

Every device

Sort it, filter it, compare a few.

Sort any column, or tick rows to compare a few head to head. Value is capability per dollar (memory × bandwidth ÷ price), so a box with lots of fast memory for the money scores high. Price is the one that matters: new while a part is still sold, used once it is discontinued. Apple, mini-PC and Lucebox memory is the full unified pool.

Device ⇅ Class ⇅ Memory GB ⇅ Bandwidth GB/s ⇅ Price ⇅ Value ⇅
★ Luceboxdetails →
Lucebox Zero 495: Radeon AI PRO R9700 32GB + up to 192GB unified, tuned Lucebox engine
Lucebox 160 u 640 $5,999 new 17
MI400 / MI455X 432GB
HBM4, first Helios rack shipments late Q3 2026; per-GPU price estimated
Datacenter 432 19,600 $50,000 new 169
B200 192GB Datacenter 192 8,000 $48,000 new 32
B300 288GB
Blackwell Ultra
Datacenter 288 8,000 $55,000 new 42
MI325X 256GB Datacenter 256 6,000 $26,000 new 59
MI300X 192GB Datacenter 192 5,300 $16,000 new 64
H200 141GB Datacenter 141 4,800 $37,000 new 18
H100 80GB SXM Datacenter 80 3,350 $32,000 new 8
A100 80GB Datacenter 80 2,039 $6,500 used 25
RTX 5090 32GB
Out of stock at US retailers; marketplace price vs $1,999 MSRP (GDDR7 shortage)
Consumer GPU 32 1,792 $6,450 new 9
RTX PRO 6000 Blackwell 96GB
96GB workstation, 1792 GB/s
Workstation 96 1,792 $15,499 new 11
RTX PRO 5000 Blackwell 72GB Workstation 72 1,344 $4,900 new 20
Mac Studio M5 Ultra 96GB
Shipping since September 2026, 1.2 TB/s
Apple 96 u 1,200 $5,499 new 21
Mac Studio M5 Ultra 256GB
512GB ships late October
Apple 256 u 1,200 $11,999 new 26
AMD Instinct MI50 32GB
Cheap-VRAM favorite, used only
Datacenter 32 1,024 $320 used 102
RTX 3090 Ti 24GB Consumer GPU 24 1,008 $1,300 used 19
RTX 4090 24GB Consumer GPU 24 1,008 $2,500 used 10
RTX 5080 16GB Consumer GPU 16 960 $1,650 new 9
RX 7900 XTX 24GB Consumer GPU 24 960 $900 used 26
RTX 6000 Ada 48GB Workstation 48 960 $6,800 new 7
RTX 3090 24GB
24GB used-market favorite for local AI
Consumer GPU 24 936 $1,350 used 17
RTX 3080 12GB Consumer GPU 12 912 $340 used 32
Intel Arc Pro B60 Dual 48GB
Two B60 GPUs on one card
Workstation 48 912 $1,200 new 36
Tesla V100 32GB Datacenter 32 900 $700 used 41
RTX 5070 Ti 16GB Consumer GPU 16 896 $1,200 new 12
Radeon PRO W7900 48GB Workstation 48 864 $3,897 new 11
Radeon PRO W7800 48GB
Gigabyte AI TOP 48GB variant
Workstation 48 864 $2,700 new 15
L40S 48GB Datacenter 48 864 $8,500 new 5
Mac Studio M3 Ultra 96GB
Replaced by the M5 Ultra Mac Studio
Apple 96 u 819 $3,300 used 24
Mac Studio M3 Ultra 512GB
512GB retired; the M5 Ultra 512GB ships late October
Apple 512 u 819 $30,000 used 14
RX 7900 XT 20GB Consumer GPU 20 800 $700 used 23
Mac Studio M2 Ultra 192GB Apple 192 u 800 $5,500 used 28
RTX A6000 48GB Workstation 48 768 $3,500 used 11
RTX 3080 10GB Consumer GPU 10 760 $370 used 21
RTX 4080 Super 16GB Consumer GPU 16 736 $1,100 used 11
RTX 4080 16GB Consumer GPU 16 717 $1,080 used 11
RTX 4070 Ti Super 16GB Consumer GPU 16 672 $770 used 14
RTX 5070 12GB Consumer GPU 12 672 $1,000 new 8
RX 9070 XT 16GB Consumer GPU 16 640 $800 new 13
RX 9070 16GB Consumer GPU 16 640 $690 new 15
Radeon AI PRO R9700 32GB Workstation 32 640 $1,800 new 11
RX 7800 XT 16GB Consumer GPU 16 624 $500 used 20
RTX 2080 Ti 22GB (modded)
China-modded 22GB refurb
Consumer GPU 22 616 $550 used 25
MacBook Pro M5 Max 128GB Apple 128 u 614 $6,749 new 12
Tenstorrent Wormhole n300 24GB
RISC-V Tensix, fully OSS stack
Workstation 24 576 $1,399 new 10
Mac Studio M4 Max 128GB
40-core; 128GB pulled, max now 96GB
Apple 128 u 546 $3,200 used 22
RX 6800 16GB Consumer GPU 16 512 $380 used 22
Tenstorrent Blackhole p150 32GB
RISC-V Blackhole, fully OSS stack
Workstation 32 512 $1,399 new 12
RTX 4070 12GB Consumer GPU 12 504 $500 used 12
Intel Arc Pro B60 24GB Workstation 24 456 $674 new 16
RTX 3060 Ti 8GB Consumer GPU 8 448 $240 used 15
Mac Studio M4 Max 36GB
32-core GPU, ~410 GB/s; replaced by the M5 Max Mac Studio
Apple 36 u 410 $2,100 used 7
RTX 3060 12GB Consumer GPU 12 360 $300 used 14
Tesla P40 24GB
Cheap 24GB, no video out, needs blower
Datacenter 24 347 $270 used 31
MacBook Pro M5 Pro 64GB Apple 64 u 307 $3,499 new 6
Mac mini M4 Pro 64GB
64GB config pulled (DRAM shortage)
Apple 64 u 273 $2,200 used 8
DGX Spark 128GB (GB10)
Coherent memory + CUDA
Mini-PC 128 u 273 $6,500 new 5
Ryzen AI Max+ 395 128GB
Strix Halo, ~96GB allocatable to GPU
Mini-PC 128 u 256 $3,499 new 9
Jetson AGX Orin 64GB
275 TOPS edge dev kit; was $1,999 before July 2026
Mini-PC 64 u 204 $3,499 new 4
MacBook Air M5 32GB Apple 32 u 153 $1,499 new 3
Mac mini M4 16GB
Replaced by the M6 Mac mini in September 2026
Apple 16 u 120 $850 used 2

Method. Memory bandwidth is peak theoretical. Each device shows the price that actually applies: the current new price while it is still sold, the current used price once it is discontinued. Figures are indicative US-market numbers triangulated across retailers (Newegg, Amazon, B&H, Apple) and used marketplaces (eBay sold listings, resellers) as of 2026-09-29; the 2026 memory shortage has pushed new prices up and pulled several high-memory configs from sale. Value = memory × bandwidth ÷ price (capability per dollar). Datacenter parts are quote-based and sold in 8-GPU boards, so their per-unit figures are approximate. "u" marks unified memory. Numbers move fast; treat them as a snapshot, not a quote.

Our build

Lucebox, full disclosure: we make this one.

A Radeon AI PRO R9700 with 32GB GDDR6 and 640 GB/s bandwidth, paired with 128GB unified memory and a tuned stack. Plug-and-play, validated, and warrantied. The numbers above are here so you can judge the configuration directly.

Next to the others

Cost, setup, speed and support, side by side.

The comparison page puts Lucebox next to cloud APIs, a DGX Spark and a Mac Studio, row by row.