Home / Value / Cheapest GPU for Code Llama 13B Instruct
Cheapest GPU for Code Llama 13B Instruct
Q4 · 8K (default)
Ranked by lowest fresh Amazon-matched board-partner SKU price among GPUs that fully fit this exact workload. This is not a speed ranking and not a “best GPU” list. Query combinations are not indexed.
To open another model, use the value explorer or the model catalog.
Requirement
- Model
- Code Llama 13B Instruct
- Quantization · context
- Q4 · 8K
- Required VRAM
- 15.2 GB
- Calculator
- v1.0.0
35 canonical GPUs fully fit. 0 of them have a fresh current price. 35 compatible GPUs have no current price.
Formula: MIN(live fresh Amazon-matched board-partner SKU price) WHERE compatibility = FITS_IN_VRAM. Sort: price ASC, advertised VRAM ASC, canonical GPU id ASC. Prices are the lowest fresh Amazon-matched board-partner SKU, not MSRP or a street-price average.
Among 0 canonical GPUs with a fresh matched Amazon offer currently available in RigForAI. Ranked 0. Unpriced compatible: 35. This is not the entire US market — only canonical GPUs with a fresh matched Amazon offer currently in RigForAI.
Cheapest currently available full-VRAM GPU
No current matched Amazon price is available. Compatible GPUs may still be listed below. Stale prices are not used to invent a cheapest winner.
Compatible — current price unavailable
These GPUs calculate as a full VRAM fit. They are not inserted into the current-price ranking because RigForAI has no fresh matched Amazon offer for them. That is not the same as too expensive or incompatible.
- NVIDIA RTX 4500 24 GB · 24 GB advertised
- NVIDIA RTX 3090 24 GB · 24 GB advertised
- NVIDIA RTX 5000 32 GB · 32 GB advertised
- NVIDIA Tesla K80 24 GB · 24 GB advertised
- NVIDIA Quadro M6000 24 GB · 24 GB advertised
- NVIDIA Tesla M10 32 GB · 32 GB advertised
- NVIDIA Tesla M40 24 GB · 24 GB advertised
- NVIDIA Quadro P6000 24 GB · 24 GB advertised
- NVIDIA Quadro GV100 32 GB · 32 GB advertised
- NVIDIA RTX 6000 24 GB · 24 GB advertised
- NVIDIA RTX 8000 48 GB · 48 GB advertised
- NVIDIA RTX A6000 48 GB · 48 GB advertised
- NVIDIA RTX A5000 24 GB · 24 GB advertised
- NVIDIA RTX A4500 20 GB · 20 GB advertised
- NVIDIA RTX 3090 Ti 24 GB · 24 GB advertised
- NVIDIA RTX 4090 24 GB · 24 GB advertised
- NVIDIA RTX A5500 24 GB · 24 GB advertised
- AMD Radeon RX 7900 XT 20 GB · 20 GB advertised
- AMD Radeon RX 7900 XTX 24 GB · 24 GB advertised
- AMD Radeon RX 7900 XT 24 GB · 24 GB advertised
- NVIDIA RTX 6000 48 GB · 48 GB advertised
- NVIDIA RTX 4000 20 GB · 20 GB advertised
- NVIDIA RTX 5090 32 GB · 32 GB advertised
- NVIDIA H100 80 GB · 80 GB advertised
- NVIDIA H100 NVL 94 GB · 94 GB advertised
- NVIDIA A100 80 GB · 80 GB advertised
- NVIDIA A100 40 GB · 40 GB advertised
- NVIDIA L40S 48 GB · 48 GB advertised
- NVIDIA L40 48 GB · 48 GB advertised
- NVIDIA L4 24 GB · 24 GB advertised
- NVIDIA Tesla V100S 32 GB · 32 GB advertised
- NVIDIA H200 141 GB · 141 GB advertised
- NVIDIA GH200 96 GB · 96 GB advertised
- NVIDIA B200 192 GB · 192 GB advertised
- NVIDIA RTX PRO 6000 Blackwell 96 GB · 96 GB advertised
Supported llama.cpp decode per current $1,000
Not ranked here: comparable published MEDIUM-or-higher-confidence llama.cpp decode predictions with a fresh price are unavailable for this workload. Raw Phase 5 estimates may still appear on GPU and comparison pages.
Open the GPU finder · Model page · Value explorer · Workstations · Servers · Buy vs rent