Home / Models / GPT-OSS 120B / NVIDIA RTX 5000 16 GB
Can GPT-OSS 120B run on NVIDIA RTX 5000 16 GB?
No. Estimated memory for GPT-OSS 120B exceeds NVIDIA RTX 5000 16 GB, including the modeled offload allowance.
Can it run?
- Can it run?
- No (calculated estimate).
- Full GPU fit?
- No
- Model
- GPT-OSS 120B · 117B
- GPU
- NVIDIA RTX 5000 16 GB · 16 GB advertised
Memory calculation
Breakdown for Q4 at 8K context, batch size 1. Calculator v1.0.0.
| Model weights | 68.6 GB |
|---|---|
| KV cache | 0.6 GB |
| Runtime reserve | 3.9 GB |
| Safety margin | 2.3 GB |
| Required (estimated) | 75.5 GB |
| GPU usable VRAM | 14.4 GB |
Result: DOES NOT FIT
Quantization table
| Quantization | 2K | 4K | 8K | 16K | 32K | 64K | 128K |
|---|---|---|---|---|---|---|---|
| BF16 | No | No | No | No | No | No | No |
| FP16 | No | No | No | No | No | No | No |
| Q3 | No | No | No | No | No | No | No |
| Q4 | No | No | No | No | No | No | No |
| Q5 | No | No | No | No | No | No | No |
| Q6 | No | No | No | No | No | No | No |
| Q8 | No | No | No | No | No | No | No |
What fits entirely in VRAM?
No calculated full-VRAM fit at the evaluated quantizations and context lengths.
What requires RAM offload?
Offload is shown only when the model does not fully fit in usable VRAM but stays within the modeled offload allowance.
Available graphics cards
These SKUs use this GPU chip. Amazon CTAs appear only for EXACT/HIGH matches. Absence of a price does not change the compatibility result above.

HP 5JH81AA graphics card NVIDIA Quadro RTX 5000 16 GB GDDR6
16 GB · Amazon EXACT
$602.99
Check price on Amazon
Fujitsu S26361-F2222-L505 graphics card NVIDIA Quadro RTX 5000 16 GB GDDR6
16 GB · Amazon unmatched
Sources
Model: official config openai/gpt-oss-120b (main). GPU/product specs: Icecat. Compatibility: RigForAI calculator v1.0.0 estimated VRAM · Q8 at 8K: DOES_NOT_FIT. Amazon: affiliate commerce match, not a spec source.