Can I Run It?

Written by Jakub Rusinowski · Last updated July 30, 2026

Pick a model family below to see exactly which GPUs it fits on, with quantization and estimated tokens/sec for each.

DeepSeek R1

Llama 3.3

Llama 3.1 Family

Nemotron 70B

Command R Family

Phi-4 Family

Qwen 2.5 Family

Gemma 2 Family

Mistral Family

Yi 1.5 Family

Granite 3.0

Llama 4

Gemma 3

Gemma 4

Llama 3.2 Family

DeepSeek V3

Qwen 3

Qwen 2.5 VL

Mistral Small 3.1

Codestral

StarCoder 2

OLMo 2

Falcon 3

InternLM 3

Aya Expanse

Qwen3-Coder

GLM-4.7 / GLM-Z1

EXAONE 3.5

Llama 3.2 Vision

Ministral

Qwen 3.5

Kimi K2.5

MiniMax M2.5

MiMo-V2-Pro

DeepSeek V3.2

Nemotron Cascade 2

GPT-oss 120B

Devstral-2

Cogito v1

GLM-5 / GLM-5.1

Qwen 3.6

DeepSeek V4

Mistral Small 4

IBM Granite 4.1

MiniMax M3

Mistral Large 3

DeepSeek V4.1

Qwen 3.7

GLM-6

Llama 4.5

Bonsai 27B

Inkling

Cosmos 3

Best LLMs by VRAM Tier → | Check Your Hardware