Skip to main content

GPU Fit Matrix

Can the RTX 3090 (used) run Gemma 4 26B A4B ?

No — this GPU does not have enough VRAM

Planning estimate: Gemma 4 26B A4B (26B total / 4B active) needs about28.5 GB of VRAM at FP8(≈4K context). The RTX 3090 (used) has24 GB.

This is a computed planning estimate (parameter count × bytes/parameter × quantization + context overhead), not a measured benchmark. See how we estimate.

VRAM needed by quantization

QuantizationEst. VRAM neededFits 24GB?
BF1654.5 GB❌ No
FP828.5 GB❌ No

No measured benchmark yet. We don't publish invented tokens/sec — when we have a first-party measured run for this pairing, it will appear here. For now this page is a VRAM planning estimate. Model the exact case in ourVRAM calculator.

Related