Written by Jakub Rusinowski · Last updated September 25, 2024
Yes, comfortably — you'll have ~10 GB of headroom running Llama 3.2 90B Vision Instruct at Q4_K_M (54 GB, ~7 tok/s (est.)).
Check price on Amazon — Apple MacBook Pro M3 Max
| VRAM | 64 GB unified memory |
| Memory Bandwidth | 400 GB/s |
| Llama 3.2 90B Vision Instruct | Q4_K_M · 54 GB · ~7 tok/s (est.) |
| Llama 3.2 11B Vision Instruct | Q4_K_M · 7.8 GB · ~51 tok/s (est.) |
| Llama 3.2 3B Instruct | Q4_K_M · 2.2 GB · ~182 tok/s (est.) |
| Llama 3.2 1B Instruct | Q4_K_M · 0.8 GB · ~400 tok/s (est.) |
At 2 hrs/day, buying (~$2,499) beats renting at $0.77/hr after about 4.5 years.
Affiliate links — we may earn a commission if you sign up, at no extra cost to you.
Cloud rates verified 2026-07 — estimates, and marketplace prices vary. Buying price is GPU MSRP only, not a full PC.
Yes, comfortably — you'll have ~10 GB of headroom running Llama 3.2 90B Vision Instruct at Q4_K_M (54 GB, ~7 tok/s (est.)).
Llama 3.2 90B Vision Instruct at Q4_K_M quantization (54 GB), estimated ~7 tokens/sec.
← Can I Run It? | Llama 3.2 Family Model Page | Apple M3 Max GPU Page | Check Your Hardware