GPU Fit Matrix
Can the Radeon Pro W7900 run gpt-oss-120b?
No — this GPU does not have enough VRAM
Planning estimate: gpt-oss-120b (120B total) needs about62.5 GB of VRAM at FP4(≈4K context). The Radeon Pro W7900 has48 GB.
This is a computed planning estimate (parameter count × bytes/parameter × quantization + context overhead), not a measured benchmark. See how we estimate.
VRAM needed by quantization
| Quantization | Est. VRAM needed | Fits 48GB? |
|---|---|---|
| BF16 | 242.5 GB | ❌ No |
| FP16 | 242.5 GB | ❌ No |
| FP8 | 122.5 GB | ❌ No |
| FP4 | 62.5 GB | ❌ No |
No measured benchmark yet. We don't publish invented tokens/sec — when we have a first-party measured run for this pairing, it will appear here. For now this page is a VRAM planning estimate. Model the exact case in ourVRAM calculator.
Related
- VRAM calculator — tweak quantization and context for this exact model
- How much VRAM do you need?
- Best GPU for local LLMs
- All GPU × model fit pages