Mac Studio, M5 Max, 128GB vs Mac Studio, M5 Ultra, 96GB
Neither box holds gpt-oss-120b at Q8_0. The largest quant that fits is Q6_K, on Mac Studio, M5 Max, 128GB.
- Fit
- Won't fit
- Won't fit
- Headroom
- −6GB
- −38GB
- Est. speed
- —
- —
- Cheapest today
- $5,099 at Apple
- $5,499 at Best Buy
- Per GB of memory
- $40
- $57
- Memory
- 128GB unified memory
- 96GB unified memory
- Usable for a model
- 120GB
- 88GB
- Memory bandwidth
- 614GB/s
- 1.2TB/s
- Storage
- 512GB
- 1TB
- Lowest recorded
- $5,099 in 3d
- $5,499 in 3d
Fit, headroom and speed are computed from each vendor’s published bandwidth and the model’s own size at 8K context. Prices are the cheapest listing we could read on each seller’s page. The arithmetic is here.
Where they actually differ
- Qwen3-235B-A22BWon't fitWon't fit
- DeepSeek-V3Won't fitWon't fit
- Llama-3.3-70BFitsFits
- gpt-oss-120bWon't fitWon't fit
- Mistral-Large-2411Won't fitWon't fit
- Llama-4-Scout-17B-16ETightWon't fit
- Qwen3-32BFitsFits
- Gemma-3-27BFitsFits
- gpt-oss-20bFitsFits
- Llama-4-Maverick-17B-128EWon't fitWon't fit
- Llama-4-BehemothWon't fitWon't fit
- Qwen3-Coder-480B-A35B-InstructWon't fitWon't fit
- Kimi-K2-InstructWon't fitWon't fit
- GLM-4.6Won't fitWon't fit
- DeepSeek-R1Won't fitWon't fit
- MiniMax-M2Won't fitWon't fit
- Qwen3-30B-A3BFitsFits
- Qwen3-4BFitsFits
- Devstral-Small-2507FitsFits
- Phi-4FitsFits
- Gemma-3n-E4BFitsFits
- Qwen3-VL-8B-InstructFitsFits
Every model we track at Q8_0. The rows where the two boxes agree are dimmed, because they are not what you are choosing between.