Can I Run It?
作者: Jakub Rusinowski · 最后更新: 2026年9月19日
Pick a model family below to see exactly which GPUs it fits on, with quantization and estimated tokens/sec for each.
BitNet b1.58
- BitNet b1.58 on NVIDIA GeForce RTX 5060 Ti 8GB
- BitNet b1.58 on NVIDIA GeForce RTX 5060
- BitNet b1.58 on NVIDIA GeForce RTX 4060
DeepSeek R1
- DeepSeek R1 on NVIDIA GeForce RTX 5090
- DeepSeek R1 on NVIDIA GeForce RTX 5080
- DeepSeek R1 on NVIDIA GeForce RTX 5070 Ti
- DeepSeek R1 on NVIDIA GeForce RTX 5070
- DeepSeek R1 on NVIDIA GeForce RTX 5060 Ti 16GB
- DeepSeek R1 on NVIDIA GeForce RTX 5060 Ti 8GB
- DeepSeek R1 on NVIDIA GeForce RTX 5060
- DeepSeek R1 on NVIDIA GeForce RTX 4090
- All DeepSeek R1 pairings (38 GPUs) →
Llama 3.3
- Llama 3.3 on NVIDIA GeForce RTX 5090
- Llama 3.3 on NVIDIA GeForce RTX 5080
- Llama 3.3 on NVIDIA GeForce RTX 5070 Ti
- Llama 3.3 on NVIDIA GeForce RTX 5070
- Llama 3.3 on NVIDIA GeForce RTX 5060 Ti 16GB
- Llama 3.3 on NVIDIA GeForce RTX 4090
- Llama 3.3 on NVIDIA GeForce RTX 4080 Super
- Llama 3.3 on NVIDIA GeForce RTX 4080
- All Llama 3.3 pairings (35 GPUs) →
Llama 3.1 Family
- Llama 3.1 Family on NVIDIA GeForce RTX 5070
- Llama 3.1 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Llama 3.1 Family on NVIDIA GeForce RTX 5060
- Llama 3.1 Family on NVIDIA GeForce RTX 4070 Ti
- Llama 3.1 Family on NVIDIA GeForce RTX 4070 Super
- Llama 3.1 Family on NVIDIA GeForce RTX 4070
- Llama 3.1 Family on NVIDIA GeForce RTX 4060
- Llama 3.1 Family on NVIDIA GeForce RTX 3080 (10GB)
- All Llama 3.1 Family pairings (14 GPUs) →
Nemotron 70B
- Nemotron 70B on NVIDIA GeForce RTX 5090
- Nemotron 70B on NVIDIA GeForce RTX 4090
- Nemotron 70B on NVIDIA GeForce RTX 3090
- Nemotron 70B on AMD Radeon RX 7900 XTX
- Nemotron 70B on AMD Radeon RX 7900 XT
- Nemotron 70B on Apple M4 Pro
- Nemotron 70B on Apple M4
- Nemotron 70B on Apple M3 Max
- All Nemotron 70B pairings (23 GPUs) →
Command R Family
- Command R Family on NVIDIA GeForce RTX 5090
- Command R Family on NVIDIA GeForce RTX 5080
- Command R Family on NVIDIA GeForce RTX 5070 Ti
- Command R Family on NVIDIA GeForce RTX 5070
- Command R Family on NVIDIA GeForce RTX 5060 Ti 16GB
- Command R Family on NVIDIA GeForce RTX 4090
- Command R Family on NVIDIA GeForce RTX 4080 Super
- Command R Family on NVIDIA GeForce RTX 4080
- All Command R Family pairings (44 GPUs) →
Phi-4 Family
- Phi-4 Family on NVIDIA GeForce RTX 5080
- Phi-4 Family on NVIDIA GeForce RTX 5070 Ti
- Phi-4 Family on NVIDIA GeForce RTX 5070
- Phi-4 Family on NVIDIA GeForce RTX 5060 Ti 16GB
- Phi-4 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Phi-4 Family on NVIDIA GeForce RTX 5060
- Phi-4 Family on NVIDIA GeForce RTX 4080 Super
- Phi-4 Family on NVIDIA GeForce RTX 4080
- All Phi-4 Family pairings (27 GPUs) →
Phi 3.5 Family
- Phi 3.5 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Phi 3.5 Family on NVIDIA GeForce RTX 5060
- Phi 3.5 Family on NVIDIA GeForce RTX 4060
Qwen 2.5 Family
- Qwen 2.5 Family on NVIDIA GeForce RTX 5090
- Qwen 2.5 Family on NVIDIA GeForce RTX 5080
- Qwen 2.5 Family on NVIDIA GeForce RTX 5070 Ti
- Qwen 2.5 Family on NVIDIA GeForce RTX 5070
- Qwen 2.5 Family on NVIDIA GeForce RTX 5060 Ti 16GB
- Qwen 2.5 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen 2.5 Family on NVIDIA GeForce RTX 5060
- Qwen 2.5 Family on NVIDIA GeForce RTX 4090
- All Qwen 2.5 Family pairings (46 GPUs) →
Gemma 2 Family
- Gemma 2 Family on NVIDIA GeForce RTX 5070
- Gemma 2 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Gemma 2 Family on NVIDIA GeForce RTX 5060
- Gemma 2 Family on NVIDIA GeForce RTX 4070 Ti
- Gemma 2 Family on NVIDIA GeForce RTX 4070 Super
- Gemma 2 Family on NVIDIA GeForce RTX 4070
- Gemma 2 Family on NVIDIA GeForce RTX 4060
- Gemma 2 Family on NVIDIA GeForce RTX 3080 (10GB)
- All Gemma 2 Family pairings (14 GPUs) →
Mistral Family
- Mistral Family on NVIDIA GeForce RTX 5090
- Mistral Family on NVIDIA GeForce RTX 5080
- Mistral Family on NVIDIA GeForce RTX 5070 Ti
- Mistral Family on NVIDIA GeForce RTX 5070
- Mistral Family on NVIDIA GeForce RTX 5060 Ti 16GB
- Mistral Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Mistral Family on NVIDIA GeForce RTX 5060
- Mistral Family on NVIDIA GeForce RTX 4090
- All Mistral Family pairings (39 GPUs) →
Yi 1.5 Family
- Yi 1.5 Family on NVIDIA GeForce RTX 5090
- Yi 1.5 Family on NVIDIA GeForce RTX 5070
- Yi 1.5 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Yi 1.5 Family on NVIDIA GeForce RTX 5060
- Yi 1.5 Family on NVIDIA GeForce RTX 4090
- Yi 1.5 Family on NVIDIA GeForce RTX 4070 Ti
- Yi 1.5 Family on NVIDIA GeForce RTX 4070 Super
- Yi 1.5 Family on NVIDIA GeForce RTX 4070
- All Yi 1.5 Family pairings (25 GPUs) →
Granite 3.0
- Granite 3.0 on NVIDIA GeForce RTX 5070
- Granite 3.0 on NVIDIA GeForce RTX 5060 Ti 8GB
- Granite 3.0 on NVIDIA GeForce RTX 5060
- Granite 3.0 on NVIDIA GeForce RTX 4070 Ti
- Granite 3.0 on NVIDIA GeForce RTX 4070 Super
- Granite 3.0 on NVIDIA GeForce RTX 4070
- Granite 3.0 on NVIDIA GeForce RTX 4060
- Granite 3.0 on NVIDIA GeForce RTX 3080 (10GB)
- All Granite 3.0 pairings (14 GPUs) →
Llama 4
- Llama 4 on NVIDIA GeForce RTX 5090
- Llama 4 on Apple M4 Max
- Llama 4 on Apple M4
- Llama 4 on Apple M3 Max
- Llama 4 on Apple M3 Pro
- Llama 4 on Apple M3 Ultra
- Llama 4 on Apple M2 Max
- Llama 4 on Apple M2 Pro
- All Llama 4 pairings (21 GPUs) →
Gemma 3
- Gemma 3 on NVIDIA GeForce RTX 5090
- Gemma 3 on NVIDIA GeForce RTX 5080
- Gemma 3 on NVIDIA GeForce RTX 5070 Ti
- Gemma 3 on NVIDIA GeForce RTX 5070
- Gemma 3 on NVIDIA GeForce RTX 5060 Ti 16GB
- Gemma 3 on NVIDIA GeForce RTX 4090
- Gemma 3 on NVIDIA GeForce RTX 4080 Super
- Gemma 3 on NVIDIA GeForce RTX 4080
- All Gemma 3 pairings (29 GPUs) →
Gemma 4
- Gemma 4 on NVIDIA GeForce RTX 5090
- Gemma 4 on NVIDIA GeForce RTX 5080
- Gemma 4 on NVIDIA GeForce RTX 5070 Ti
- Gemma 4 on NVIDIA GeForce RTX 5070
- Gemma 4 on NVIDIA GeForce RTX 5060 Ti 16GB
- Gemma 4 on NVIDIA GeForce RTX 5060 Ti 8GB
- Gemma 4 on NVIDIA GeForce RTX 5060
- Gemma 4 on NVIDIA GeForce RTX 4090
- All Gemma 4 pairings (39 GPUs) →
Llama 3.2 Family
- Llama 3.2 Family on NVIDIA GeForce RTX 5070
- Llama 3.2 Family on NVIDIA GeForce RTX 5060 Ti 8GB
- Llama 3.2 Family on NVIDIA GeForce RTX 5060
- Llama 3.2 Family on NVIDIA GeForce RTX 4070 Ti
- Llama 3.2 Family on NVIDIA GeForce RTX 4070 Super
- Llama 3.2 Family on NVIDIA GeForce RTX 4070
- Llama 3.2 Family on NVIDIA GeForce RTX 4060
- Llama 3.2 Family on NVIDIA GeForce RTX 3080 (10GB)
- All Llama 3.2 Family pairings (23 GPUs) →
DeepSeek V3
Qwen 3
- Qwen 3 on NVIDIA GeForce RTX 5090
- Qwen 3 on NVIDIA GeForce RTX 5080
- Qwen 3 on NVIDIA GeForce RTX 5070 Ti
- Qwen 3 on NVIDIA GeForce RTX 5070
- Qwen 3 on NVIDIA GeForce RTX 5060 Ti 16GB
- Qwen 3 on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen 3 on NVIDIA GeForce RTX 5060
- Qwen 3 on NVIDIA GeForce RTX 4090
- All Qwen 3 pairings (39 GPUs) →
Qwen 2.5 VL
- Qwen 2.5 VL on NVIDIA GeForce RTX 5070
- Qwen 2.5 VL on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen 2.5 VL on NVIDIA GeForce RTX 5060
- Qwen 2.5 VL on NVIDIA GeForce RTX 4070 Ti
- Qwen 2.5 VL on NVIDIA GeForce RTX 4070 Super
- Qwen 2.5 VL on NVIDIA GeForce RTX 4070
- Qwen 2.5 VL on NVIDIA GeForce RTX 4060
- Qwen 2.5 VL on NVIDIA GeForce RTX 3080 (10GB)
- All Qwen 2.5 VL pairings (24 GPUs) →
Mistral Small 3.1
- Mistral Small 3.1 on NVIDIA GeForce RTX 5090
- Mistral Small 3.1 on NVIDIA GeForce RTX 5080
- Mistral Small 3.1 on NVIDIA GeForce RTX 5070 Ti
- Mistral Small 3.1 on NVIDIA GeForce RTX 5070
- Mistral Small 3.1 on NVIDIA GeForce RTX 5060 Ti 16GB
- Mistral Small 3.1 on NVIDIA GeForce RTX 5060 Ti 8GB
- Mistral Small 3.1 on NVIDIA GeForce RTX 5060
- Mistral Small 3.1 on NVIDIA GeForce RTX 4090
- All Mistral Small 3.1 pairings (38 GPUs) →
Codestral
- Codestral on NVIDIA GeForce RTX 5090
- Codestral on NVIDIA GeForce RTX 5080
- Codestral on NVIDIA GeForce RTX 5070 Ti
- Codestral on NVIDIA GeForce RTX 5070
- Codestral on NVIDIA GeForce RTX 5060 Ti 16GB
- Codestral on NVIDIA GeForce RTX 5060 Ti 8GB
- Codestral on NVIDIA GeForce RTX 5060
- Codestral on NVIDIA GeForce RTX 4090
- All Codestral pairings (38 GPUs) →
Phi-4 Mini
- Phi-4 Mini on NVIDIA GeForce RTX 5060 Ti 8GB
- Phi-4 Mini on NVIDIA GeForce RTX 5060
- Phi-4 Mini on NVIDIA GeForce RTX 4060
StarCoder 2
- StarCoder 2 on NVIDIA GeForce RTX 5080
- StarCoder 2 on NVIDIA GeForce RTX 5070 Ti
- StarCoder 2 on NVIDIA GeForce RTX 5070
- StarCoder 2 on NVIDIA GeForce RTX 5060 Ti 16GB
- StarCoder 2 on NVIDIA GeForce RTX 5060 Ti 8GB
- StarCoder 2 on NVIDIA GeForce RTX 5060
- StarCoder 2 on NVIDIA GeForce RTX 4080 Super
- StarCoder 2 on NVIDIA GeForce RTX 4080
- All StarCoder 2 pairings (27 GPUs) →
OLMo 2
- OLMo 2 on NVIDIA GeForce RTX 5080
- OLMo 2 on NVIDIA GeForce RTX 5070 Ti
- OLMo 2 on NVIDIA GeForce RTX 5060 Ti 16GB
- OLMo 2 on NVIDIA GeForce RTX 5060 Ti 8GB
- OLMo 2 on NVIDIA GeForce RTX 5060
- OLMo 2 on NVIDIA GeForce RTX 4080 Super
- OLMo 2 on NVIDIA GeForce RTX 4080
- OLMo 2 on NVIDIA GeForce RTX 4070 Ti Super
- All OLMo 2 pairings (20 GPUs) →
SmolLM2
- SmolLM2 on NVIDIA GeForce RTX 5060 Ti 8GB
- SmolLM2 on NVIDIA GeForce RTX 5060
- SmolLM2 on NVIDIA GeForce RTX 4060
Falcon 3
- Falcon 3 on NVIDIA GeForce RTX 5070
- Falcon 3 on NVIDIA GeForce RTX 5060 Ti 8GB
- Falcon 3 on NVIDIA GeForce RTX 5060
- Falcon 3 on NVIDIA GeForce RTX 4070 Ti
- Falcon 3 on NVIDIA GeForce RTX 4070 Super
- Falcon 3 on NVIDIA GeForce RTX 4070
- Falcon 3 on NVIDIA GeForce RTX 4060
- Falcon 3 on NVIDIA GeForce RTX 3080 (10GB)
- All Falcon 3 pairings (14 GPUs) →
InternLM 3
- InternLM 3 on NVIDIA GeForce RTX 5080
- InternLM 3 on NVIDIA GeForce RTX 5070 Ti
- InternLM 3 on NVIDIA GeForce RTX 5070
- InternLM 3 on NVIDIA GeForce RTX 5060 Ti 16GB
- InternLM 3 on NVIDIA GeForce RTX 5060 Ti 8GB
- InternLM 3 on NVIDIA GeForce RTX 5060
- InternLM 3 on NVIDIA GeForce RTX 4090
- InternLM 3 on NVIDIA GeForce RTX 4080 Super
- All InternLM 3 pairings (32 GPUs) →
Aya Expanse
- Aya Expanse on NVIDIA GeForce RTX 5090
- Aya Expanse on NVIDIA GeForce RTX 5070
- Aya Expanse on NVIDIA GeForce RTX 5060 Ti 8GB
- Aya Expanse on NVIDIA GeForce RTX 5060
- Aya Expanse on NVIDIA GeForce RTX 4090
- Aya Expanse on NVIDIA GeForce RTX 4070 Ti
- Aya Expanse on NVIDIA GeForce RTX 4070 Super
- Aya Expanse on NVIDIA GeForce RTX 4070
- All Aya Expanse pairings (25 GPUs) →
Qwen3-Coder
- Qwen3-Coder on NVIDIA GeForce RTX 5070
- Qwen3-Coder on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen3-Coder on NVIDIA GeForce RTX 5060
- Qwen3-Coder on NVIDIA GeForce RTX 4070 Ti
- Qwen3-Coder on NVIDIA GeForce RTX 4070 Super
- Qwen3-Coder on NVIDIA GeForce RTX 4070
- Qwen3-Coder on NVIDIA GeForce RTX 4060
- Qwen3-Coder on NVIDIA GeForce RTX 3080 (10GB)
- All Qwen3-Coder pairings (20 GPUs) →
GLM-4.7 / GLM-Z1
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 5090
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 5070
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 5060 Ti 8GB
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 5060
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 4090
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 4070 Ti
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 4070 Super
- GLM-4.7 / GLM-Z1 on NVIDIA GeForce RTX 4070
- All GLM-4.7 / GLM-Z1 pairings (26 GPUs) →
Aya 3B (Tiny Aya)
- Aya 3B (Tiny Aya) on NVIDIA GeForce RTX 5060 Ti 8GB
- Aya 3B (Tiny Aya) on NVIDIA GeForce RTX 5060
- Aya 3B (Tiny Aya) on NVIDIA GeForce RTX 4060
EXAONE 3.5
- EXAONE 3.5 on NVIDIA GeForce RTX 5090
- EXAONE 3.5 on NVIDIA GeForce RTX 5060 Ti 8GB
- EXAONE 3.5 on NVIDIA GeForce RTX 5060
- EXAONE 3.5 on NVIDIA GeForce RTX 4090
- EXAONE 3.5 on NVIDIA GeForce RTX 4060
- EXAONE 3.5 on NVIDIA GeForce RTX 3090
- EXAONE 3.5 on NVIDIA GeForce RTX 3080 (10GB)
- EXAONE 3.5 on NVIDIA GeForce RTX 3070 Ti
- All EXAONE 3.5 pairings (19 GPUs) →
Gemma 3n
- Gemma 3n on NVIDIA GeForce RTX 5060 Ti 8GB
- Gemma 3n on NVIDIA GeForce RTX 5060
- Gemma 3n on NVIDIA GeForce RTX 4060
- Gemma 3n on NVIDIA GeForce RTX 3080 (10GB)
- Gemma 3n on NVIDIA GeForce RTX 3070 Ti
- Gemma 3n on NVIDIA GeForce RTX 3070
- Gemma 3n on AMD Radeon RX 9060 XT 8GB
- Gemma 3n on Intel Arc B570
Llama 3.2 Vision
- Llama 3.2 Vision on NVIDIA GeForce RTX 5070
- Llama 3.2 Vision on NVIDIA GeForce RTX 5060 Ti 8GB
- Llama 3.2 Vision on NVIDIA GeForce RTX 5060
- Llama 3.2 Vision on NVIDIA GeForce RTX 4070 Ti
- Llama 3.2 Vision on NVIDIA GeForce RTX 4070 Super
- Llama 3.2 Vision on NVIDIA GeForce RTX 4070
- Llama 3.2 Vision on NVIDIA GeForce RTX 4060
- Llama 3.2 Vision on NVIDIA GeForce RTX 3080 (10GB)
- All Llama 3.2 Vision pairings (23 GPUs) →
Ministral
- Ministral on NVIDIA GeForce RTX 5070
- Ministral on NVIDIA GeForce RTX 5060 Ti 8GB
- Ministral on NVIDIA GeForce RTX 5060
- Ministral on NVIDIA GeForce RTX 4070 Ti
- Ministral on NVIDIA GeForce RTX 4070 Super
- Ministral on NVIDIA GeForce RTX 4070
- Ministral on NVIDIA GeForce RTX 4060
- Ministral on NVIDIA GeForce RTX 3080 (10GB)
- All Ministral pairings (14 GPUs) →
Qwen 3.5
- Qwen 3.5 on NVIDIA GeForce RTX 5090
- Qwen 3.5 on NVIDIA GeForce RTX 5070
- Qwen 3.5 on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen 3.5 on NVIDIA GeForce RTX 5060
- Qwen 3.5 on NVIDIA GeForce RTX 4090
- Qwen 3.5 on NVIDIA GeForce RTX 4070 Ti
- Qwen 3.5 on NVIDIA GeForce RTX 4070 Super
- Qwen 3.5 on NVIDIA GeForce RTX 4070
- All Qwen 3.5 pairings (38 GPUs) →
Kimi K2.5
MiniMax M2.5
- MiniMax M2.5 on Apple M4 Max
- MiniMax M2.5 on Apple M3 Max
- MiniMax M2.5 on Apple M2 Ultra
- MiniMax M2.5 on Apple M2 Max
- MiniMax M2.5 on Apple M1 Ultra
- MiniMax M2.5 on Apple M1 Max
- MiniMax M2.5 on Apple M5 Pro
- MiniMax M2.5 on Apple M5 Max
- All MiniMax M2.5 pairings (13 GPUs) →
MiMo-V2-Pro
DeepSeek V3.2
Nemotron Cascade 2
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5090
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5080
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5070 Ti
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5070
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5060 Ti 16GB
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5060 Ti 8GB
- Nemotron Cascade 2 on NVIDIA GeForce RTX 5060
- Nemotron Cascade 2 on NVIDIA GeForce RTX 4090
- All Nemotron Cascade 2 pairings (49 GPUs) →
GPT-OSS
- GPT-OSS on NVIDIA GeForce RTX 5080
- GPT-OSS on NVIDIA GeForce RTX 5070 Ti
- GPT-OSS on NVIDIA GeForce RTX 5070
- GPT-OSS on NVIDIA GeForce RTX 5060 Ti 16GB
- GPT-OSS on NVIDIA GeForce RTX 5060 Ti 8GB
- GPT-OSS on NVIDIA GeForce RTX 5060
- GPT-OSS on NVIDIA GeForce RTX 4090
- GPT-OSS on NVIDIA GeForce RTX 4080 Super
- All GPT-OSS pairings (42 GPUs) →
Devstral
- Devstral on NVIDIA GeForce RTX 5090
- Devstral on NVIDIA GeForce RTX 5080
- Devstral on NVIDIA GeForce RTX 5070 Ti
- Devstral on NVIDIA GeForce RTX 5070
- Devstral on NVIDIA GeForce RTX 5060 Ti 16GB
- Devstral on NVIDIA GeForce RTX 5060 Ti 8GB
- Devstral on NVIDIA GeForce RTX 5060
- Devstral on NVIDIA GeForce RTX 4090
- All Devstral pairings (47 GPUs) →
Cogito v1
- Cogito v1 on NVIDIA GeForce RTX 5090
- Cogito v1 on NVIDIA GeForce RTX 5080
- Cogito v1 on NVIDIA GeForce RTX 5070 Ti
- Cogito v1 on NVIDIA GeForce RTX 5070
- Cogito v1 on NVIDIA GeForce RTX 5060 Ti 16GB
- Cogito v1 on NVIDIA GeForce RTX 5060 Ti 8GB
- Cogito v1 on NVIDIA GeForce RTX 5060
- Cogito v1 on NVIDIA GeForce RTX 4090
- All Cogito v1 pairings (46 GPUs) →
Kimi K3
GLM-5 / GLM-5.1
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 5090
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 5070
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 5060 Ti 8GB
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 5060
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 4090
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 4070 Ti
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 4070 Super
- GLM-5 / GLM-5.1 on NVIDIA GeForce RTX 4070
- All GLM-5 / GLM-5.1 pairings (33 GPUs) →
Qwen 3.6
- Qwen 3.6 on NVIDIA GeForce RTX 5090
- Qwen 3.6 on NVIDIA GeForce RTX 5080
- Qwen 3.6 on NVIDIA GeForce RTX 5070 Ti
- Qwen 3.6 on NVIDIA GeForce RTX 5070
- Qwen 3.6 on NVIDIA GeForce RTX 5060 Ti 16GB
- Qwen 3.6 on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen 3.6 on NVIDIA GeForce RTX 5060
- Qwen 3.6 on NVIDIA GeForce RTX 4090
- All Qwen 3.6 pairings (41 GPUs) →
DeepSeek V4
- DeepSeek V4 on Apple M4 Max
- DeepSeek V4 on Apple M2 Ultra
- DeepSeek V4 on Apple M2 Max
- DeepSeek V4 on Apple M1 Ultra
- DeepSeek V4 on Apple M5 Max
- DeepSeek V4 on NVIDIA DGX Spark
- DeepSeek V4 on AMD Ryzen AI Max+ 395
- DeepSeek V4 on NVIDIA RTX PRO 6000 Blackwell
- All DeepSeek V4 pairings (10 GPUs) →
Mistral Small 4
- Mistral Small 4 on NVIDIA GeForce RTX 5090
- Mistral Small 4 on Apple M4 Max
- Mistral Small 4 on Apple M4
- Mistral Small 4 on Apple M3 Max
- Mistral Small 4 on Apple M3 Pro
- Mistral Small 4 on Apple M2 Max
- Mistral Small 4 on Apple M2 Pro
- Mistral Small 4 on Apple M1 Ultra
- All Mistral Small 4 pairings (20 GPUs) →
IBM Granite 4.1
- IBM Granite 4.1 on NVIDIA GeForce RTX 5070
- IBM Granite 4.1 on NVIDIA GeForce RTX 5060 Ti 8GB
- IBM Granite 4.1 on NVIDIA GeForce RTX 5060
- IBM Granite 4.1 on NVIDIA GeForce RTX 4070 Ti
- IBM Granite 4.1 on NVIDIA GeForce RTX 4070 Super
- IBM Granite 4.1 on NVIDIA GeForce RTX 4070
- IBM Granite 4.1 on NVIDIA GeForce RTX 4060
- IBM Granite 4.1 on NVIDIA GeForce RTX 3080 (10GB)
- All IBM Granite 4.1 pairings (19 GPUs) →
MiniMax M3
- MiniMax M3 on Apple M4 Max
- MiniMax M3 on Apple M3 Max
- MiniMax M3 on Apple M2 Ultra
- MiniMax M3 on Apple M2 Max
- MiniMax M3 on Apple M1 Ultra
- MiniMax M3 on Apple M1 Max
- MiniMax M3 on Apple M5 Pro
- MiniMax M3 on Apple M5 Max
- All MiniMax M3 pairings (13 GPUs) →
Mistral Large 3
DeepSeek V4.1
- DeepSeek V4.1 on Apple M4 Max
- DeepSeek V4.1 on Apple M2 Ultra
- DeepSeek V4.1 on Apple M2 Max
- DeepSeek V4.1 on Apple M1 Ultra
- DeepSeek V4.1 on Apple M5 Max
- DeepSeek V4.1 on NVIDIA DGX Spark
- DeepSeek V4.1 on AMD Ryzen AI Max+ 395
- DeepSeek V4.1 on NVIDIA RTX PRO 6000 Blackwell
- All DeepSeek V4.1 pairings (10 GPUs) →
Qwen 3.7
- Qwen 3.7 on NVIDIA GeForce RTX 5090
- Qwen 3.7 on NVIDIA GeForce RTX 5080
- Qwen 3.7 on NVIDIA GeForce RTX 5070 Ti
- Qwen 3.7 on NVIDIA GeForce RTX 5070
- Qwen 3.7 on NVIDIA GeForce RTX 5060 Ti 16GB
- Qwen 3.7 on NVIDIA GeForce RTX 4090
- Qwen 3.7 on NVIDIA GeForce RTX 4080 Super
- Qwen 3.7 on NVIDIA GeForce RTX 4080
- All Qwen 3.7 pairings (35 GPUs) →
GLM-6
- GLM-6 on NVIDIA GeForce RTX 5070
- GLM-6 on NVIDIA GeForce RTX 5060 Ti 8GB
- GLM-6 on NVIDIA GeForce RTX 5060
- GLM-6 on NVIDIA GeForce RTX 4070 Ti
- GLM-6 on NVIDIA GeForce RTX 4070 Super
- GLM-6 on NVIDIA GeForce RTX 4070
- GLM-6 on NVIDIA GeForce RTX 4060
- GLM-6 on NVIDIA GeForce RTX 3080 (10GB)
- All GLM-6 pairings (15 GPUs) →
Llama 4.5
- Llama 4.5 on NVIDIA GeForce RTX 5090
- Llama 4.5 on Apple M4 Max
- Llama 4.5 on Apple M4
- Llama 4.5 on Apple M3 Max
- Llama 4.5 on Apple M3 Pro
- Llama 4.5 on Apple M2 Max
- Llama 4.5 on Apple M2 Pro
- Llama 4.5 on Apple M1 Ultra
- All Llama 4.5 pairings (20 GPUs) →
Bonsai 27B
- Bonsai 27B on NVIDIA GeForce RTX 5070
- Bonsai 27B on NVIDIA GeForce RTX 5060 Ti 8GB
- Bonsai 27B on NVIDIA GeForce RTX 5060
- Bonsai 27B on NVIDIA GeForce RTX 4070 Ti
- Bonsai 27B on NVIDIA GeForce RTX 4070 Super
- Bonsai 27B on NVIDIA GeForce RTX 4070
- Bonsai 27B on NVIDIA GeForce RTX 4060
- Bonsai 27B on NVIDIA GeForce RTX 3080 (10GB)
- All Bonsai 27B pairings (14 GPUs) →
Inkling
Cosmos 3
- Cosmos 3 on NVIDIA GeForce RTX 5080
- Cosmos 3 on NVIDIA GeForce RTX 5070 Ti
- Cosmos 3 on NVIDIA GeForce RTX 5070
- Cosmos 3 on NVIDIA GeForce RTX 5060 Ti 16GB
- Cosmos 3 on NVIDIA GeForce RTX 5060 Ti 8GB
- Cosmos 3 on NVIDIA GeForce RTX 5060
- Cosmos 3 on NVIDIA GeForce RTX 4080 Super
- Cosmos 3 on NVIDIA GeForce RTX 4080
- All Cosmos 3 pairings (35 GPUs) →
Poolside Laguna XS 2.1
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5090
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5080
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5070 Ti
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5070
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5060 Ti 16GB
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5060 Ti 8GB
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 5060
- Poolside Laguna XS 2.1 on NVIDIA GeForce RTX 4090
- All Poolside Laguna XS 2.1 pairings (41 GPUs) →
Nemotron 3 Super
- Nemotron 3 Super on NVIDIA GeForce RTX 5090
- Nemotron 3 Super on Apple M4 Max
- Nemotron 3 Super on Apple M4
- Nemotron 3 Super on Apple M3 Max
- Nemotron 3 Super on Apple M3 Pro
- Nemotron 3 Super on Apple M2 Max
- Nemotron 3 Super on Apple M2 Pro
- Nemotron 3 Super on Apple M1 Ultra
- All Nemotron 3 Super pairings (20 GPUs) →
GLM-5.2
MiniMax M2.7
- MiniMax M2.7 on Apple M4 Max
- MiniMax M2.7 on Apple M3 Max
- MiniMax M2.7 on Apple M2 Ultra
- MiniMax M2.7 on Apple M2 Max
- MiniMax M2.7 on Apple M1 Ultra
- MiniMax M2.7 on Apple M1 Max
- MiniMax M2.7 on Apple M5 Pro
- MiniMax M2.7 on Apple M5 Max
- All MiniMax M2.7 pairings (13 GPUs) →
IBM Granite 4.0
- IBM Granite 4.0 on NVIDIA GeForce RTX 5090
- IBM Granite 4.0 on NVIDIA GeForce RTX 5080
- IBM Granite 4.0 on NVIDIA GeForce RTX 5070 Ti
- IBM Granite 4.0 on NVIDIA GeForce RTX 5070
- IBM Granite 4.0 on NVIDIA GeForce RTX 5060 Ti 16GB
- IBM Granite 4.0 on NVIDIA GeForce RTX 5060 Ti 8GB
- IBM Granite 4.0 on NVIDIA GeForce RTX 5060
- IBM Granite 4.0 on NVIDIA GeForce RTX 4090
- All IBM Granite 4.0 pairings (41 GPUs) →
SmolLM3
- SmolLM3 on NVIDIA GeForce RTX 5060 Ti 8GB
- SmolLM3 on NVIDIA GeForce RTX 5060
- SmolLM3 on NVIDIA GeForce RTX 4060
VibeThinker
- VibeThinker on NVIDIA GeForce RTX 5060 Ti 8GB
- VibeThinker on NVIDIA GeForce RTX 5060
- VibeThinker on NVIDIA GeForce RTX 4060
Magistral Small
- Magistral Small on NVIDIA GeForce RTX 5090
- Magistral Small on NVIDIA GeForce RTX 5080
- Magistral Small on NVIDIA GeForce RTX 5070 Ti
- Magistral Small on NVIDIA GeForce RTX 5070
- Magistral Small on NVIDIA GeForce RTX 5060 Ti 16GB
- Magistral Small on NVIDIA GeForce RTX 5060 Ti 8GB
- Magistral Small on NVIDIA GeForce RTX 5060
- Magistral Small on NVIDIA GeForce RTX 4090
- All Magistral Small pairings (39 GPUs) →
Mistral Small 3.2
- Mistral Small 3.2 on NVIDIA GeForce RTX 5090
- Mistral Small 3.2 on NVIDIA GeForce RTX 5080
- Mistral Small 3.2 on NVIDIA GeForce RTX 5070 Ti
- Mistral Small 3.2 on NVIDIA GeForce RTX 5070
- Mistral Small 3.2 on NVIDIA GeForce RTX 5060 Ti 16GB
- Mistral Small 3.2 on NVIDIA GeForce RTX 5060 Ti 8GB
- Mistral Small 3.2 on NVIDIA GeForce RTX 5060
- Mistral Small 3.2 on NVIDIA GeForce RTX 4090
- All Mistral Small 3.2 pairings (38 GPUs) →
Nemotron 3 Nano Omni
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5090
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5080
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5070 Ti
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5070
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5060 Ti 16GB
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5060 Ti 8GB
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 5060
- Nemotron 3 Nano Omni on NVIDIA GeForce RTX 4090
- All Nemotron 3 Nano Omni pairings (39 GPUs) →
Qwen3.8
- Qwen3.8 on NVIDIA GeForce RTX 5090
- Qwen3.8 on NVIDIA GeForce RTX 5080
- Qwen3.8 on NVIDIA GeForce RTX 5070 Ti
- Qwen3.8 on NVIDIA GeForce RTX 5070
- Qwen3.8 on NVIDIA GeForce RTX 5060 Ti 16GB
- Qwen3.8 on NVIDIA GeForce RTX 5060 Ti 8GB
- Qwen3.8 on NVIDIA GeForce RTX 5060
- Qwen3.8 on NVIDIA GeForce RTX 4090
- All Qwen3.8 pairings (41 GPUs) →
MiMo-V2.5
- MiMo-V2.5 on Apple M4 Max
- MiMo-V2.5 on Apple M2 Ultra
- MiMo-V2.5 on Apple M2 Max
- MiMo-V2.5 on Apple M1 Ultra
- MiMo-V2.5 on Apple M5 Max
- MiMo-V2.5 on NVIDIA DGX Spark
- MiMo-V2.5 on AMD Ryzen AI Max+ 395
- MiMo-V2.5 on NVIDIA RTX PRO 6000 Blackwell
- All MiMo-V2.5 pairings (10 GPUs) →
Ministral 3
- Ministral 3 on NVIDIA GeForce RTX 5080
- Ministral 3 on NVIDIA GeForce RTX 5070 Ti
- Ministral 3 on NVIDIA GeForce RTX 5070
- Ministral 3 on NVIDIA GeForce RTX 5060 Ti 16GB
- Ministral 3 on NVIDIA GeForce RTX 5060 Ti 8GB
- Ministral 3 on NVIDIA GeForce RTX 5060
- Ministral 3 on NVIDIA GeForce RTX 4080 Super
- Ministral 3 on NVIDIA GeForce RTX 4080
- All Ministral 3 pairings (27 GPUs) →
Nex-N2.5
- Nex-N2.5 on NVIDIA GeForce RTX 5090
- Nex-N2.5 on NVIDIA GeForce RTX 5080
- Nex-N2.5 on NVIDIA GeForce RTX 5070 Ti
- Nex-N2.5 on NVIDIA GeForce RTX 5070
- Nex-N2.5 on NVIDIA GeForce RTX 5060 Ti 16GB
- Nex-N2.5 on NVIDIA GeForce RTX 4090
- Nex-N2.5 on NVIDIA GeForce RTX 4080 Super
- Nex-N2.5 on NVIDIA GeForce RTX 4080
- All Nex-N2.5 pairings (36 GPUs) →
Nex-N2
- Nex-N2 on NVIDIA GeForce RTX 5090
- Nex-N2 on NVIDIA GeForce RTX 5080
- Nex-N2 on NVIDIA GeForce RTX 5070 Ti
- Nex-N2 on NVIDIA GeForce RTX 5070
- Nex-N2 on NVIDIA GeForce RTX 5060 Ti 16GB
- Nex-N2 on NVIDIA GeForce RTX 4090
- Nex-N2 on NVIDIA GeForce RTX 4080 Super
- Nex-N2 on NVIDIA GeForce RTX 4080
- All Nex-N2 pairings (36 GPUs) →
Nex-N1
GLM-5.3-Flash
- GLM-5.3-Flash on Apple M4 Max
- GLM-5.3-Flash on Apple M2 Ultra
- GLM-5.3-Flash on Apple M2 Max
- GLM-5.3-Flash on Apple M1 Ultra
- GLM-5.3-Flash on Apple M5 Max
- GLM-5.3-Flash on NVIDIA DGX Spark
- GLM-5.3-Flash on AMD Ryzen AI Max+ 395
- GLM-5.3-Flash on NVIDIA RTX PRO 6000 Blackwell
- All GLM-5.3-Flash pairings (10 GPUs) →
North Mini Code
- North Mini Code on NVIDIA GeForce RTX 5090
- North Mini Code on NVIDIA GeForce RTX 5080
- North Mini Code on NVIDIA GeForce RTX 5070 Ti
- North Mini Code on NVIDIA GeForce RTX 5070
- North Mini Code on NVIDIA GeForce RTX 5060 Ti 16GB
- North Mini Code on NVIDIA GeForce RTX 5060 Ti 8GB
- North Mini Code on NVIDIA GeForce RTX 5060
- North Mini Code on NVIDIA GeForce RTX 4090
- All North Mini Code pairings (39 GPUs) →
IBM Granite 4.2
- IBM Granite 4.2 on NVIDIA GeForce RTX 5090
- IBM Granite 4.2 on NVIDIA GeForce RTX 5070
- IBM Granite 4.2 on NVIDIA GeForce RTX 5060 Ti 8GB
- IBM Granite 4.2 on NVIDIA GeForce RTX 5060
- IBM Granite 4.2 on NVIDIA GeForce RTX 4090
- IBM Granite 4.2 on NVIDIA GeForce RTX 4070 Ti
- IBM Granite 4.2 on NVIDIA GeForce RTX 4070 Super
- IBM Granite 4.2 on NVIDIA GeForce RTX 4070
- All IBM Granite 4.2 pairings (24 GPUs) →
MiniCPM-V
- MiniCPM-V on NVIDIA GeForce RTX 5070
- MiniCPM-V on NVIDIA GeForce RTX 5060 Ti 8GB
- MiniCPM-V on NVIDIA GeForce RTX 5060
- MiniCPM-V on NVIDIA GeForce RTX 4070 Ti
- MiniCPM-V on NVIDIA GeForce RTX 4070 Super
- MiniCPM-V on NVIDIA GeForce RTX 4070
- MiniCPM-V on NVIDIA GeForce RTX 4060
- MiniCPM-V on NVIDIA GeForce RTX 3080 (10GB)
- All MiniCPM-V pairings (14 GPUs) →
DeepSeek-OCR
- DeepSeek-OCR on NVIDIA GeForce RTX 5060 Ti 8GB
- DeepSeek-OCR on NVIDIA GeForce RTX 5060
- DeepSeek-OCR on NVIDIA GeForce RTX 4060
LFM2.5
- LFM2.5 on NVIDIA GeForce RTX 5070
- LFM2.5 on NVIDIA GeForce RTX 5060 Ti 8GB
- LFM2.5 on NVIDIA GeForce RTX 5060
- LFM2.5 on NVIDIA GeForce RTX 4070 Ti
- LFM2.5 on NVIDIA GeForce RTX 4070 Super
- LFM2.5 on NVIDIA GeForce RTX 4070
- LFM2.5 on NVIDIA GeForce RTX 4060
- LFM2.5 on NVIDIA GeForce RTX 3080 (10GB)
- All LFM2.5 pairings (14 GPUs) →
By system RAM and by computer
Not everyone is asking about a graphics card. These check the same models against a system-RAM budget for CPU-only inference, or against a specific machine.
8 GB system RAM
- BitNet b1.58 on 8 GB RAM
- DeepSeek R1 on 8 GB RAM
- Llama 3.1 Family on 8 GB RAM
- Phi 3.5 Family on 8 GB RAM
- Qwen 2.5 Family on 8 GB RAM
- Mistral Family on 8 GB RAM
16 GB system RAM
- BitNet b1.58 on 16 GB RAM
- DeepSeek R1 on 16 GB RAM
- Llama 3.1 Family on 16 GB RAM
- Phi-4 Family on 16 GB RAM
- Phi 3.5 Family on 16 GB RAM
- Qwen 2.5 Family on 16 GB RAM
24 GB system RAM
- BitNet b1.58 on 24 GB RAM
- DeepSeek R1 on 24 GB RAM
- Llama 3.1 Family on 24 GB RAM
- Phi-4 Family on 24 GB RAM
- Phi 3.5 Family on 24 GB RAM
- Qwen 2.5 Family on 24 GB RAM
32 GB system RAM
- BitNet b1.58 on 32 GB RAM
- DeepSeek R1 on 32 GB RAM
- Llama 3.1 Family on 32 GB RAM
- Command R Family on 32 GB RAM
- Phi-4 Family on 32 GB RAM
- Phi 3.5 Family on 32 GB RAM
48 GB system RAM
- BitNet b1.58 on 48 GB RAM
- DeepSeek R1 on 48 GB RAM
- Llama 3.3 on 48 GB RAM
- Llama 3.1 Family on 48 GB RAM
- Nemotron 70B on 48 GB RAM
- Command R Family on 48 GB RAM
64 GB system RAM
- BitNet b1.58 on 64 GB RAM
- DeepSeek R1 on 64 GB RAM
- Llama 3.3 on 64 GB RAM
- Llama 3.1 Family on 64 GB RAM
- Nemotron 70B on 64 GB RAM
- Command R Family on 64 GB RAM
96 GB system RAM
- BitNet b1.58 on 96 GB RAM
- DeepSeek R1 on 96 GB RAM
- Llama 3.3 on 96 GB RAM
- Llama 3.1 Family on 96 GB RAM
- Nemotron 70B on 96 GB RAM
- Command R Family on 96 GB RAM
128 GB system RAM
- BitNet b1.58 on 128 GB RAM
- DeepSeek R1 on 128 GB RAM
- Llama 3.3 on 128 GB RAM
- Llama 3.1 Family on 128 GB RAM
- Nemotron 70B on 128 GB RAM
- Command R Family on 128 GB RAM
192 GB system RAM
- BitNet b1.58 on 192 GB RAM
- DeepSeek R1 on 192 GB RAM
- Llama 3.3 on 192 GB RAM
- Llama 3.1 Family on 192 GB RAM
- Nemotron 70B on 192 GB RAM
- Command R Family on 192 GB RAM
256 GB system RAM
- BitNet b1.58 on 256 GB RAM
- DeepSeek R1 on 256 GB RAM
- Llama 3.3 on 256 GB RAM
- Llama 3.1 Family on 256 GB RAM
- Nemotron 70B on 256 GB RAM
- Command R Family on 256 GB RAM
MacBook Pro 16" (M4 Max, 128 GB)
- BitNet b1.58 on MacBook Pro M4 Max 128 GB
- DeepSeek R1 on MacBook Pro M4 Max 128 GB
- Llama 3.3 on MacBook Pro M4 Max 128 GB
- Llama 3.1 Family on MacBook Pro M4 Max 128 GB
- Nemotron 70B on MacBook Pro M4 Max 128 GB
- Command R Family on MacBook Pro M4 Max 128 GB
MacBook Pro 16" (M4 Max, 48 GB)
- BitNet b1.58 on MacBook Pro M4 Max 48 GB
- DeepSeek R1 on MacBook Pro M4 Max 48 GB
- Llama 3.3 on MacBook Pro M4 Max 48 GB
- Llama 3.1 Family on MacBook Pro M4 Max 48 GB
- Nemotron 70B on MacBook Pro M4 Max 48 GB
- Command R Family on MacBook Pro M4 Max 48 GB
MacBook Pro 14" (M4 Pro, 24 GB)
- BitNet b1.58 on MacBook Pro M4 Pro 24 GB
- DeepSeek R1 on MacBook Pro M4 Pro 24 GB
- Llama 3.1 Family on MacBook Pro M4 Pro 24 GB
- Phi-4 Family on MacBook Pro M4 Pro 24 GB
- Phi 3.5 Family on MacBook Pro M4 Pro 24 GB
- Qwen 2.5 Family on MacBook Pro M4 Pro 24 GB
MacBook Air (M4, 16 GB)
- BitNet b1.58 on MacBook Air M4 16 GB
- DeepSeek R1 on MacBook Air M4 16 GB
- Llama 3.1 Family on MacBook Air M4 16 GB
- Phi-4 Family on MacBook Air M4 16 GB
- Phi 3.5 Family on MacBook Air M4 16 GB
- Qwen 2.5 Family on MacBook Air M4 16 GB
MacBook Pro 16" (M5 Max, 128 GB)
- BitNet b1.58 on MacBook Pro M5 Max 128 GB
- DeepSeek R1 on MacBook Pro M5 Max 128 GB
- Llama 3.3 on MacBook Pro M5 Max 128 GB
- Llama 3.1 Family on MacBook Pro M5 Max 128 GB
- Nemotron 70B on MacBook Pro M5 Max 128 GB
- Command R Family on MacBook Pro M5 Max 128 GB
Mac Studio (M3 Ultra, 256 GB)
- BitNet b1.58 on Mac Studio M3 Ultra 256 GB
- DeepSeek R1 on Mac Studio M3 Ultra 256 GB
- Llama 3.3 on Mac Studio M3 Ultra 256 GB
- Llama 3.1 Family on Mac Studio M3 Ultra 256 GB
- Nemotron 70B on Mac Studio M3 Ultra 256 GB
- Command R Family on Mac Studio M3 Ultra 256 GB
Mac Studio (M4 Max, 64 GB)
- BitNet b1.58 on Mac Studio M4 Max 64 GB
- DeepSeek R1 on Mac Studio M4 Max 64 GB
- Llama 3.3 on Mac Studio M4 Max 64 GB
- Llama 3.1 Family on Mac Studio M4 Max 64 GB
- Nemotron 70B on Mac Studio M4 Max 64 GB
- Command R Family on Mac Studio M4 Max 64 GB
Mac mini (M4, 16 GB)
- BitNet b1.58 on Mac mini M4 16 GB
- DeepSeek R1 on Mac mini M4 16 GB
- Llama 3.1 Family on Mac mini M4 16 GB
- Phi-4 Family on Mac mini M4 16 GB
- Phi 3.5 Family on Mac mini M4 16 GB
- Qwen 2.5 Family on Mac mini M4 16 GB
Mac mini (M4 Pro, 64 GB)
- BitNet b1.58 on Mac mini M4 Pro 64 GB
- DeepSeek R1 on Mac mini M4 Pro 64 GB
- Llama 3.3 on Mac mini M4 Pro 64 GB
- Llama 3.1 Family on Mac mini M4 Pro 64 GB
- Nemotron 70B on Mac mini M4 Pro 64 GB
- Command R Family on Mac mini M4 Pro 64 GB
Mac Studio (M2 Ultra, 192 GB)
- BitNet b1.58 on Mac Studio M2 Ultra 192 GB
- DeepSeek R1 on Mac Studio M2 Ultra 192 GB
- Llama 3.3 on Mac Studio M2 Ultra 192 GB
- Llama 3.1 Family on Mac Studio M2 Ultra 192 GB
- Nemotron 70B on Mac Studio M2 Ultra 192 GB
- Command R Family on Mac Studio M2 Ultra 192 GB
Framework Desktop (Ryzen AI Max+ 395, 128 GB)
- BitNet b1.58 on Framework Desktop 128 GB
- DeepSeek R1 on Framework Desktop 128 GB
- Llama 3.3 on Framework Desktop 128 GB
- Llama 3.1 Family on Framework Desktop 128 GB
- Nemotron 70B on Framework Desktop 128 GB
- Command R Family on Framework Desktop 128 GB
Beelink SER9 (Ryzen AI 9, 32 GB)
- BitNet b1.58 on Beelink SER9 32 GB
- DeepSeek R1 on Beelink SER9 32 GB
- Llama 3.3 on Beelink SER9 32 GB
- Llama 3.1 Family on Beelink SER9 32 GB
- Nemotron 70B on Beelink SER9 32 GB
- Command R Family on Beelink SER9 32 GB
RTX 5090 Desktop (32 GB VRAM, 64 GB RAM)
- BitNet b1.58 on RTX 5090 desktop
- DeepSeek R1 on RTX 5090 desktop
- Llama 3.3 on RTX 5090 desktop
- Llama 3.1 Family on RTX 5090 desktop
- Nemotron 70B on RTX 5090 desktop
- Command R Family on RTX 5090 desktop
RTX 4090 Desktop (24 GB VRAM, 64 GB RAM)
- BitNet b1.58 on RTX 4090 desktop
- DeepSeek R1 on RTX 4090 desktop
- Llama 3.1 Family on RTX 4090 desktop
- Command R Family on RTX 4090 desktop
- Phi-4 Family on RTX 4090 desktop
- Phi 3.5 Family on RTX 4090 desktop
RTX 3090 Desktop (24 GB VRAM, 64 GB RAM)
- BitNet b1.58 on RTX 3090 desktop
- DeepSeek R1 on RTX 3090 desktop
- Llama 3.1 Family on RTX 3090 desktop
- Command R Family on RTX 3090 desktop
- Phi-4 Family on RTX 3090 desktop
- Phi 3.5 Family on RTX 3090 desktop
RTX 5080 Desktop (16 GB VRAM, 32 GB RAM)
- BitNet b1.58 on RTX 5080 desktop
- DeepSeek R1 on RTX 5080 desktop
- Llama 3.1 Family on RTX 5080 desktop
- Phi-4 Family on RTX 5080 desktop
- Phi 3.5 Family on RTX 5080 desktop
- Qwen 2.5 Family on RTX 5080 desktop
RTX 5060 Ti 16 GB Desktop (16 GB VRAM, 32 GB RAM)
- BitNet b1.58 on RTX 5060 Ti 16 GB desktop
- DeepSeek R1 on RTX 5060 Ti 16 GB desktop
- Llama 3.1 Family on RTX 5060 Ti 16 GB desktop
- Phi-4 Family on RTX 5060 Ti 16 GB desktop
- Phi 3.5 Family on RTX 5060 Ti 16 GB desktop
- Qwen 2.5 Family on RTX 5060 Ti 16 GB desktop
RTX 3060 12 GB Desktop (12 GB VRAM, 32 GB RAM)
- BitNet b1.58 on RTX 3060 12 GB desktop
- DeepSeek R1 on RTX 3060 12 GB desktop
- Llama 3.1 Family on RTX 3060 12 GB desktop
- Phi-4 Family on RTX 3060 12 GB desktop
- Phi 3.5 Family on RTX 3060 12 GB desktop
- Qwen 2.5 Family on RTX 3060 12 GB desktop
RTX 4090 Laptop (16 GB VRAM, 32 GB RAM)
- BitNet b1.58 on RTX 4090 laptop
- DeepSeek R1 on RTX 4090 laptop
- Llama 3.1 Family on RTX 4090 laptop
- Phi-4 Family on RTX 4090 laptop
- Phi 3.5 Family on RTX 4090 laptop
- Qwen 2.5 Family on RTX 4090 laptop