Home / Value / Cheapest GPU for Mixtral 8x7B Instruct
Cheapest GPU for Mixtral 8x7B Instruct
Q4 · 8K (default)
Ranked by lowest fresh Amazon-matched board-partner SKU price among GPUs that fully fit this exact workload. This is not a speed ranking and not a “best GPU” list. Query combinations are not indexed.
To open another model, use the value explorer or the model catalog.
Requirement
- Model
- Mixtral 8x7B Instruct
- Quantization · context
- Q4 · 8K
- Required VRAM
- 31.3 GB
- Calculator
- v1.0.0
3 canonical GPUs fully fit. 1 of them have a fresh current price. 2 compatible GPUs have no current price.
Formula: MIN(live fresh Amazon-matched board-partner SKU price) WHERE compatibility = FITS_IN_VRAM. Sort: price ASC, advertised VRAM ASC, canonical GPU id ASC. Prices are the lowest fresh Amazon-matched board-partner SKU, not MSRP or a street-price average.
Among 1 canonical GPUs with a fresh matched Amazon offer currently available in RigForAI. Ranked 1. Unpriced compatible: 2. This is not the entire US market — only canonical GPUs with a fresh matched Amazon offer currently in RigForAI.
Cheapest currently available full-VRAM GPU
NVIDIA RTX 6000 48 GB
- Advertised VRAM
- 48 GB
- Usable VRAM
- 43.2 GB
- Headroom
- 11.9 GB
- Compatibility
- FITS_IN_VRAM
- Current fresh Amazon price
- $20888.00
- Winning physical SKU
- HP NVIDIA RTX 6000 Ada 48 GB 4DP GraphicsHPASIN B0CTP2HHBN
- Offer fetched
- 2026-08-14 20:19 UTC
Why it ranks first
- FITS_IN_VRAM for this model, quantization, and context
- current fresh matched Amazon offer (EXACT/HIGH, not a bundle)
- lowest price among 1 eligible priced GPUs
- price ASC, advertised VRAM ASC, canonical GPU id ASC.
Build a complete PC around this option. GPU-only price is not a complete-build cost.
Compatible — current price unavailable
These GPUs calculate as a full VRAM fit. They are not inserted into the current-price ranking because RigForAI has no fresh matched Amazon offer for them. That is not the same as too expensive or incompatible.
- NVIDIA RTX 8000 48 GB · 48 GB advertised
- NVIDIA RTX A6000 48 GB · 48 GB advertised
Supported llama.cpp decode per current $1,000
Not ranked here: comparable published MEDIUM-or-higher-confidence llama.cpp decode predictions with a fresh price are unavailable for this workload. Raw Phase 5 estimates may still appear on GPU and comparison pages.