Home / Models

Calculated memory fit

AI models

Each page answers how much GPU memory a model needs, and which catalog GPUs can hold it at a given quantization and context length. Figures are estimates from official model configs, not measured benchmarks.

Find a GPU

DeepSeek

ModelPublisherParametersNative contextArchitecture
DeepSeek R1 Distill Qwen 14BDeepSeek15B131,072qwen2
DeepSeek R1 Distill Qwen 32BDeepSeek33B131,072qwen2
DeepSeek LLM 7B ChatDeepSeek6.9B4,096llama
DeepSeek R1 Distill Llama 70BDeepSeek71B131,072llama
DeepSeek R1 Distill Qwen 7BDeepSeek7.6B131,072qwen2
DeepSeek R1 Distill Llama 8BDeepSeek8.0B131,072llama

Falcon

ModelPublisherParametersNative contextArchitecture
Falcon 3 10B InstructTII10B32,768llama
Falcon 7B InstructTII7.2Bfalcon
Falcon 3 7B InstructTII7.5B32,768llama

Gemma

ModelPublisherParametersNative contextArchitecture
Gemma 2 2B InstructGoogle2.6B8,192gemma2
Gemma 2 27B InstructGoogle27B8,192gemma2
Gemma 2 9B InstructGoogle9.2B8,192gemma2

GLM

ModelPublisherParametersNative contextArchitecture
GLM-4-9B-0414Zhipu AI9.4B32,768glm4
GLM-4 9B ChatZhipu AI9.4Bchatglm

GPT-OSS

ModelPublisherParametersNative contextArchitecture
GPT-OSS 120BOpenAI117B131,072gpt_oss
GPT-OSS 20BOpenAI22B131,072gpt_oss

Hermes

ModelPublisherParametersNative contextArchitecture
Hermes 3 Llama 3.1 70BNousResearch71B131,072llama
Hermes 3 Llama 3.1 8BNousResearch8.0B131,072llama

Llama

ModelPublisherParametersNative contextArchitecture
Llama 3.2 1B InstructMeta1.2B131,072llama
Code Llama 13B InstructMeta13B16,384llama
Llama 3.2 3B InstructMeta3.2B131,072llama
Code Llama 34B InstructMeta34B16,384llama
Code Llama 7B InstructMeta6.7B16,384llama
Llama 3.1 70B InstructMeta71B131,072llama
Llama 3.3 70B InstructMeta71B131,072llama
Llama 3.1 8B InstructMeta8.0B131,072llama
Llama 3 8B InstructMeta8.0B8,192llama

Mistral

ModelPublisherParametersNative contextArchitecture
Mistral Nemo 12B InstructMistral12B131,072mistral
Mistral Small 24B InstructMistral24B32,768mistral
Mistral 7B Instruct v0.3Mistral7.3B32,768mistral

Mixtral

ModelPublisherParametersNative contextArchitecture
Mixtral 8x22B InstructMistral141B65,536mixtral
Mixtral 8x7B InstructMistral47B32,768mixtral

OLMo

ModelPublisherParametersNative contextArchitecture
OLMo 2 13B InstructAllenAI14B4,096olmo2
OLMo 2 7B InstructAllenAI7.3B4,096olmo2

OpenChat

ModelPublisherParametersNative contextArchitecture
OpenChat 3.5 7BOpenChat7.2B8,192mistral

Phi

ModelPublisherParametersNative contextArchitecture
Phi-3 Medium 128K InstructMicrosoft14B131,072phi3
Phi-3 Medium 4K InstructMicrosoft14B4,096phi3
Phi-4Microsoft15B16,384phi3
Phi-3.5 Mini InstructMicrosoft3.8B131,072phi3
Phi-3 Mini 128K InstructMicrosoft3.8B131,072phi3
Phi-3 Mini 4K InstructMicrosoft3.8B4,096phi3
Phi-4 Mini InstructMicrosoft3.8B131,072phi3

Qwen

ModelPublisherParametersNative contextArchitecture
Qwen2.5 14B InstructAlibaba15B32,768qwen2
Qwen3 14BAlibaba15B40,960qwen3
Qwen2.5 1.5B InstructQwen1.5B32,768qwen2
Qwen2.5-Coder 1.5B InstructQwen1.5B32,768qwen2
Qwen2.5 3B InstructAlibaba3.1B32,768qwen2
Qwen2.5 32B InstructAlibaba33B32,768qwen2
Qwen2.5-Coder 32B InstructAlibaba33B32,768qwen2
QwQ 32BAlibaba33B40,960qwen2
Qwen3 32BAlibaba33B40,960qwen3
Qwen2.5 0.5B InstructQwen494M32,768qwen2
Qwen2.5 72B InstructAlibaba73B32,768qwen2
Qwen2.5 7B InstructAlibaba7.6B32,768qwen2
Qwen2.5-Coder 7B InstructAlibaba7.6B32,768qwen2
Qwen2.5-Math 7B InstructQwen7.6B4,096qwen2
Qwen2 7B InstructQwen7.6B32,768qwen2
Qwen3 8BAlibaba8.2B40,960qwen3

SmolLM

ModelPublisherParametersNative contextArchitecture
SmolLM2 1.7B InstructHuggingFaceTB1.7B8,192llama

Solar

ModelPublisherParametersNative contextArchitecture
SOLAR 10.7B InstructUpstage11B4,096llama

StarCoder

ModelPublisherParametersNative contextArchitecture
StarCoder2 15BBigCode16B16,384starcoder2
StarCoder2 7BBigCode7.2B16,384starcoder2

TinyLlama

ModelPublisherParametersNative contextArchitecture
TinyLlama 1.1B Chat v1.0TinyLlama1.1B2,048llama

Vicuna

ModelPublisherParametersNative contextArchitecture
Vicuna 13B v1.5LMSYS13B4,096llama
Vicuna 7B v1.5LMSYS6.7B4,096llama

Yi

ModelPublisherParametersNative contextArchitecture
Yi 1.5 34B Chat01.AI34B4,096llama
Yi 1.5 9B Chat01.AI8.8B4,096llama