Skip to main content

GPU Fit Matrix

Can the Mac Studio M4 Max (128GB) run Qwen3 235B A22B Instruct 2507?

No — this GPU does not have enough VRAM

Planning estimate: Qwen3 235B A22B Instruct 2507 (235B total / 22B active) needs about237.5 GB of VRAM at FP8(≈4K context). The Mac Studio M4 Max (128GB) has128 GB.

This is a computed planning estimate (parameter count × bytes/parameter × quantization + context overhead), not a measured benchmark. See how we estimate.

VRAM needed by quantization

QuantizationEst. VRAM neededFits 128GB?
BF16472.5 GB❌ No
FP16472.5 GB❌ No
FP8237.5 GB❌ No

No measured benchmark yet. We don't publish invented tokens/sec — when we have a first-party measured run for this pairing, it will appear here. For now this page is a VRAM planning estimate. Model the exact case in ourVRAM calculator.

Related