Home / Servers

Rent · GPU / dedicated AI compute · not generic hosting

Find a server for your AI model

Memory-capable means the selected model, quantization and context fit the server's GPU memory (and system RAM) using the same calculators as local hardware. Serving numbers, when shown, are official vLLM nightly measurements for matching GPU identity — not Phase 5 single-user tok/s and not this exact rental host.

Compare buy vs rent using these rental prices against a current local build or workstation.

Serving evaluation uses measured vLLM TTFT and TPOT at an explicit input/output shape and concurrency. Missing evidence is reported as insufficient, not as a fail.

Workload shapes are only those present in official vLLM nightly evidence. Request rate is not interchangeable with concurrency. P95 is not published by this source. Targets are yours, not an industry SLA.

Memory filter: Qwen3 235B-A22B at Q4 / 8K. Offers that are memory-capable remain listed even when serving evidence is absent.

ConfigurationProviderLocationPriceStock
2× NVIDIA H100 NVL 94 GB
503 GB RAM · 4930 GB SSD
Vast.aiUS$5.335/hour · ~$3894/month equivalent (hourly × 730)Listed as availableProvider

Priced, stock not confirmed as available

Shown separately. Not ranked as currently available.

ConfigurationProviderLocationPriceStock
4× NVIDIA L40S 48 GB
384 GB RAM · ? NVME
OVHcloudUS$5355/month · $4355 setupStock not confirmed by providerProvider
4× NVIDIA L40S 48 GB
384 GB RAM · ? NVME
OVHcloudEU$5355/month · $4355 setupStock not confirmed by providerProvider
4× NVIDIA L40S 48 GB
384 GB RAM · ? NVME
OVHcloudCA$5355/month · $4355 setupStock not confirmed by providerProvider
4× NVIDIA A100 80 GB
1024 GB RAM · 2 × 450 GB SSD
Vultr$7000/month · $0 setupNot listed as availableProvider

Providers

GPU / dedicated AI compute only. Not a generic hosting directory. All providers with inventory

Want to own the hardware? Build a PC · Prefer a prebuilt? Workstations