Which AI can your machine run?

Find out which AI models your machine can actually run.

Based on detected hardware

of 83 models run well

GPU | VRAM
Your machine Detecting hardware…

Each bar is the memory a model needs at Q4_K_M quantization, measured on the same ruler across the page. Bars stopping before the orange line run comfortably. Bars crossing it need to borrow system RAM, which is much slower.

Llama 3.1 8B Meta · 8B · Chat, Coding, Reasoning · ctx 128K · Dense · Llama 3.1 Community · 2 years ago
Qwen 3.5 9B Alibaba · 9B · Chat, Vision · ctx 32K · Dense · Apache 2.0 · 6 months ago
Ornith 1.0 9B DeepReinforce · 9B · Coding, Reasoning · ctx 256K · Dense · MIT · 2 months ago
Phi-4 14B Microsoft · 14B · Reasoning, Coding · ctx 16K · Dense · MIT · 1 year ago
GPT-OSS 20B OpenAI · 21B (active 4B) · Chat, Reasoning, Coding · ctx 128K · MoE · Apache 2.0 · 1 year ago
Mistral Small 3.1 24B Mistral AI · 24B · Chat, Vision, Coding · ctx 128K · Dense · Apache 2.0 · 1 year ago
Gemma 3 27B Google · 27B · Chat, Vision, Reasoning · ctx 128K · Dense · Gemma · 1 year ago
Qwen 2.5 Coder 32B Alibaba · 32B · Coding · ctx 128K · Dense · Apache 2.0 · 1 year ago
Qwen 3 32B Alibaba · 32B · Chat, Coding, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
DeepSeek R1 Distill 32B DeepSeek · 32B · Reasoning · ctx 64K · Dense · MIT · 1 year ago
Ornith 1.0 35B-A3B DeepReinforce · 35B (active 3B) · Coding, Reasoning · ctx 256K · MoE · MIT · 2 months ago
Llama 3.3 70B Meta · 70B · Chat, Reasoning, Coding · ctx 128K · Dense · Llama 3.3 Community · 1 year ago
Llama 4 Scout 17B Meta · 109B (active 17B) · Chat, Vision, Reasoning · ctx 128K · MoE · Llama 4 Community · 1 year ago
GPT-OSS 120B OpenAI · 117B (active 5B) · Chat, Reasoning, Coding · ctx 128K · MoE · Apache 2.0 · 1 year ago
Devstral 2 123B Mistral AI · 123B · Coding · ctx 256K · Dense · MRL · 8 months ago
DeepSeek R1 DeepSeek · 671B (active 37B) · Reasoning · ctx 64K · MoE · MIT · 1 year ago
DeepSeek V3.2 DeepSeek · 685B (active 37B) · Chat, Coding, Reasoning · ctx 128K · MoE · MIT · 8 months ago
GLM-5.2 Zhipu AI · 753B (active 40B) · Chat, Reasoning, Coding · ctx 1M · MoE · MIT · 2 months ago
Kimi K2 Moonshot AI · 1T (active 32B) · Chat, Reasoning, Coding · ctx 128K · MoE · Kimi · 1 year ago
All models
Qwen 3 0.6B Alibaba · 0.6B · Chat, Edge · ctx 32K · Dense · Apache 2.0 · 1 year ago
Qwen 3.5 0.8B Alibaba · 0.8B · Chat, Edge · ctx 32K · Dense · Apache 2.0 · 6 months ago
Llama 3.2 1B Meta · 1B · Chat, Edge · ctx 128K · Dense · Llama 3.2 Community · 1 year ago
Gemma 3 1B Google · 1B · Chat, Edge · ctx 32K · Dense · Gemma · 1 year ago
TinyLlama 1.1B Community · 1.1B · Chat, Edge · ctx 2K · Dense · Apache 2.0 · 2 years ago
Qwen 2.5 Coder 1.5B Alibaba · 1.5B · Coding · ctx 32K · Dense · Apache 2.0 · 1 year ago
DeepSeek R1 1.5B DeepSeek · 1.5B · Reasoning · ctx 64K · Dense · MIT · 1 year ago
Qwen 3 1.7B Alibaba · 1.7B · Chat, Multilingual · ctx 32K · Dense · Apache 2.0 · 1 year ago
Qwen 3.5 2B Alibaba · 2B · Chat, Multilingual · ctx 32K · Dense · Apache 2.0 · 6 months ago
Gemma 2 2B Google · 2B · Chat, Edge · ctx 8K · Dense · Gemma · 2 years ago
Llama 3.2 3B Meta · 3B · Chat, Coding · ctx 128K · Dense · Llama 3.2 Community · 1 year ago
SmolLM3 3B HuggingFace · 3B · Chat, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
Phi-3.5 Mini Microsoft · 3.8B · Reasoning, Coding, Chat · ctx 128K · Dense · MIT · 2 years ago
Phi-4 Mini Reasoning Microsoft · 3.8B · Reasoning · ctx 16K · Dense · MIT · 1 year ago
Qwen 3 4B Alibaba · 4B · Chat, Coding · ctx 32K · Dense · Apache 2.0 · 1 year ago
Gemma 3 4B Google · 4B · Chat, Vision · ctx 128K · Dense · Gemma · 1 year ago
Qwen 3.5 4B Alibaba · 4B · Chat, Multilingual · ctx 32K · Dense · Apache 2.0 · 6 months ago
Gemma 4 E2B IT Google · 5B · Chat, Vision · ctx 256K · Dense · Gemma · 4 months ago
Gemma 4 E2B Google · 5B · Vision · ctx 256K · Dense · Gemma · 4 months ago
Mistral 7B v0.3 Mistral AI · 7B · Chat, Reasoning · ctx 32K · Dense · Apache 2.0 · 2 years ago
Qwen 2.5 7B Alibaba · 7B · Chat, Multilingual, Coding · ctx 128K · Dense · Apache 2.0 · 1 year ago
Qwen 2.5 Coder 7B Alibaba · 7B · Coding · ctx 128K · Dense · Apache 2.0 · 1 year ago
DeepSeek R1 Distill 7B DeepSeek · 7B · Reasoning · ctx 64K · Dense · MIT · 1 year ago
Gemma 4 E4B IT Google · 8B · Chat, Vision · ctx 256K · Dense · Gemma · 4 months ago
Gemma 4 E4B Google · 8B · Vision · ctx 256K · Dense · Gemma · 4 months ago
Qwen 3 8B Alibaba · 8B · Chat, Coding, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
Ministral 8B Mistral AI · 8B · Chat · ctx 32K · Dense · MRL · 1 year ago
Gemma 2 9B Google · 9B · Chat, Reasoning · ctx 8K · Dense · Gemma · 2 years ago
GLM-4 9B Zhipu AI · 9B · Chat, Multilingual, Coding · ctx 128K · Dense · GLM-4 · 2 years ago
Nemotron Nano 9B v2 NVIDIA · 9B · Reasoning · ctx 128K · Dense · NVIDIA Open · 1 year ago
Llama 3.2 11B Vision Meta · 11B · Chat, Vision · ctx 128K · Dense · Llama 3.2 Community · 1 year ago
Gemma 3 12B Google · 12B · Chat, Vision, Reasoning · ctx 128K · Dense · Gemma · 1 year ago
Mistral Nemo 12B Mistral AI · 12B · Chat, Multilingual · ctx 128K · Dense · Apache 2.0 · 2 years ago
Qwen 2.5 14B Alibaba · 14B · Chat, Multilingual, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
Qwen 3 14B Alibaba · 14B · Chat, Coding, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
DeepSeek R1 Distill 14B DeepSeek · 14B · Reasoning · ctx 64K · Dense · MIT · 1 year ago
LFM2 24B Liquid AI · 24B (active 2B) · Chat, Edge, RAG · ctx 32K · MoE · Liquid AI · 9 months ago
Devstral Small 2 24B Mistral AI · 24B · Coding · ctx 256K · Dense · Apache 2.0 · 8 months ago
DiffusionGemma 26B-A4B IT Google · 26B (active 4B) · Chat, Vision, Reasoning · ctx 256K · MoE · Apache 2.0 · 2 months ago
Gemma 2 27B Google · 27B · Chat, Reasoning · ctx 8K · Dense · Gemma · 2 years ago
Gemma 4 26B-A4B IT Google · 27B (active 4B) · Chat, Vision, Reasoning · ctx 256K · MoE · Gemma · 4 months ago
Gemma 4 26B-A4B Google · 27B (active 4B) · Vision, Reasoning · ctx 256K · MoE · Gemma · 4 months ago
Qwen 3.5 27B Alibaba · 27.8B · Chat, Vision, Reasoning · ctx 256K · Dense · Apache 2.0 · 6 months ago
Qwen 3 30B-A3B Alibaba · 30B (active 3B) · Chat, Reasoning · ctx 128K · MoE · Apache 2.0 · 1 year ago
Nemotron 3 Nano 30B NVIDIA · 30B (active 3B) · Chat, Reasoning · ctx 1M · MoE · NVIDIA Open · 1 year ago
Qwen 2.5 32B Alibaba · 32B · Chat, Multilingual, Reasoning · ctx 128K · Dense · Apache 2.0 · 1 year ago
EXAONE 4.0 32B LG AI · 32B · Chat, Reasoning · ctx 128K · Dense · EXAONE AI · 1 year ago
OLMo 2 32B Allen AI · 32B · Chat, Reasoning · ctx 4K · Dense · Apache 2.0 · 1 year ago
Gemma 4 31B IT Google · 33B · Chat, Vision, Reasoning · ctx 256K · Dense · Gemma · 4 months ago
Gemma 4 31B Google · 33B · Vision, Reasoning · ctx 256K · Dense · Gemma · 4 months ago
Command R 35B Cohere · 35B · Chat, RAG · ctx 128K · Dense · CC BY-NC 4.0 · 2 years ago
Qwen 3.5 35B-A3B Alibaba · 35B (active 3B) · Chat, Vision · ctx 256K · MoE · Apache 2.0 · 6 months ago
Mixtral 8x7B Mistral AI · 47B (active 13B) · Chat, Coding · ctx 32K · MoE · Apache 2.0 · 2 years ago
Qwen 2.5 72B Alibaba · 72B · Chat, Multilingual, Reasoning, Coding · ctx 128K · Dense · Qwen · 1 year ago
Qwen 3.5 122B-A10B Alibaba · 122B (active 10B) · Chat, Vision, Reasoning · ctx 256K · MoE · Apache 2.0 · 6 months ago
Mixtral 8x22B Mistral AI · 141B (active 39B) · Chat, Coding, Reasoning · ctx 64K · MoE · Apache 2.0 · 2 years ago
Qwen 3 235B-A22B Alibaba · 235B (active 22B) · Chat, Coding, Reasoning · ctx 128K · MoE · Apache 2.0 · 1 year ago
Qwen 3.5 397B-A17B Alibaba · 397B (active 17B) · Chat, Vision, Reasoning, Coding · ctx 256K · MoE · Apache 2.0 · 6 months ago
Llama 4 Maverick 17B-128E Meta · 400B (active 17B) · Chat, Vision, Reasoning, Coding · ctx 1M · MoE · Llama 4 Community · 1 year ago
Llama 3.1 405B Meta · 405B · Chat, Reasoning, Coding · ctx 128K · Dense · Llama 3.1 Community · 2 years ago
Qwen 3 Coder 480B Alibaba · 480B (active 35B) · Coding · ctx 256K · MoE · Apache 2.0 · 1 year ago
DeepSeek V3.1 DeepSeek · 671B (active 37B) · Chat, Coding, Reasoning · ctx 128K · MoE · MIT · 1 year ago
GLM-5 Zhipu AI · 744B (active 40B) · Chat, Reasoning, Coding · ctx 128K · MoE · MIT · 6 months ago
GLM-5.1 Zhipu AI · 754B (active 40B) · Chat, Reasoning, Coding · ctx 128K · MoE · MIT · 4 months ago