Can I Run It?
Written by Jakub Rusinowski · Last updated July 30, 2026
Pick a model family below to see exactly which GPUs it fits on, with quantization and estimated tokens/sec for each.
DeepSeek R1
Llama 3.3
Llama 3.1 Family
Nemotron 70B
Command R Family
Phi-4 Family
Qwen 2.5 Family
Gemma 2 Family
Mistral Family
Yi 1.5 Family
Granite 3.0
Llama 4
Gemma 3
Gemma 4
Llama 3.2 Family
DeepSeek V3
Qwen 3
Qwen 2.5 VL
Mistral Small 3.1
Codestral
StarCoder 2
OLMo 2
Falcon 3
InternLM 3
Aya Expanse
Qwen3-Coder
GLM-4.7 / GLM-Z1
EXAONE 3.5
Llama 3.2 Vision
Ministral
Qwen 3.5
Kimi K2.5
MiniMax M2.5
MiMo-V2-Pro
DeepSeek V3.2
Nemotron Cascade 2
GPT-oss 120B
Devstral-2
Cogito v1
GLM-5 / GLM-5.1
Qwen 3.6
DeepSeek V4
Mistral Small 4
IBM Granite 4.1
MiniMax M3
Mistral Large 3
DeepSeek V4.1
Qwen 3.7
GLM-6
Llama 4.5
Bonsai 27B
Inkling
Cosmos 3
Best LLMs by VRAM Tier → | Check Your Hardware