Home / Value / Cheapest GPU for Llama 3.3 70B Instruct
Cheapest GPU for Llama 3.3 70B Instruct
Q4 · 8K (default)
Ranked by lowest fresh Amazon-matched board-partner SKU price among GPUs that fully fit this exact workload. This is not a speed ranking and not a “best GPU” list. Query combinations are not indexed.
To open another model, use the value explorer or the model catalog.
Requirement
- Model
- Llama 3.3 70B Instruct
- Quantization · context
- Q4 · 8K
- Required VRAM
- —
- Calculator
- v1.0.0
0 canonical GPUs fully fit. 0 of them have a fresh current price. 0 compatible GPUs have no current price.
Formula: MIN(live fresh Amazon-matched board-partner SKU price) WHERE compatibility = FITS_IN_VRAM. Sort: price ASC, advertised VRAM ASC, canonical GPU id ASC. Prices are the lowest fresh Amazon-matched board-partner SKU, not MSRP or a street-price average.
Among 0 canonical GPUs with a fresh matched Amazon offer currently available in RigForAI. Ranked 0. Unpriced compatible: 0. This is not the entire US market — only canonical GPUs with a fresh matched Amazon offer currently in RigForAI.
Cheapest currently available full-VRAM GPU
No current matched Amazon price is available. Compatible GPUs may still be listed below. Stale prices are not used to invent a cheapest winner. Explore multi-GPU options.
Supported llama.cpp decode per current $1,000
Not ranked here: comparable published MEDIUM-or-higher-confidence llama.cpp decode predictions with a fresh price are unavailable for this workload. Raw Phase 5 estimates may still appear on GPU and comparison pages.