Local AI Deals

DGX Spark, GB10 Grace Blackwell, 128GB vs Mac Studio, M5 Max, 128GB

Both hold Qwen3-235B-A22B at Q3_K_M. Mac Studio, M5 Max, 128GB decodes an est. 28.6 tok/s against 12.7, and costs $400 more — $25.22.580645161290704 per extra token per second. Worth it if you read the output as it arrives; not if you queue jobs.

For Qwen3-235B-A22B at Q3_K_MDGX Spark 128GBMac Studio 128GB
Fit
Tight
Tight
Headroom
3GB
2GB
Est. speed
12.7 tok/s
28.6 tok/s
Cheapest today
$4,699 at NVIDIA
$5,099 at Apple
Per GB of memory
$37
$40
Memory
128GB unified LPDDR5X memory
128GB unified memory
Usable for a model
120GB
120GB
Memory bandwidth
273GB/s
614GB/s
Storage
4TB
512GB
Lowest recorded
$4,699 in 2d
$5,099 in 2d

Fit, headroom and speed are computed from each vendor’s published bandwidth and the model’s own size at 8K context. Prices are the cheapest listing we could read on each seller’s page. The arithmetic is here.

Where they actually differ

Every model we track at Q3_K_M. The rows where the two boxes agree are dimmed, because they are not what you are choosing between.