Autor: Jakub Rusinowski · Ostatnia aktualizacja: 25 września 2024
Yes, comfortably — you'll have ~8.2 GB of headroom running Llama 3.2 11B Vision Instruct at Q4_K_M (7.8 GB, ~123 tok/s (est.)).
Sprawdź cenę na Amazon — NVIDIA GeForce RTX 5080 16GB
| VRAM | 16 GB |
| Memory Bandwidth | 960 GB/s |
| Llama 3.2 11B Vision Instruct | Q4_K_M · 7.8 GB · ~123 tok/s (est.) |
| Llama 3.2 3B Instruct | Q4_K_M · 2.2 GB · ~400 tok/s (est.) |
| Llama 3.2 1B Instruct | Q4_K_M · 0.8 GB · ~400 tok/s (est.) |
At 2 hrs/day, buying (~$999) beats renting at $0.34/hr after about 4.1 years.
Affiliate links — we may earn a commission if you sign up, at no extra cost to you.
Cloud rates verified 2026-07 — estimates, and marketplace prices vary. Buying price is GPU MSRP only, not a full PC.
Yes, comfortably — you'll have ~8.2 GB of headroom running Llama 3.2 11B Vision Instruct at Q4_K_M (7.8 GB, ~123 tok/s (est.)).
Llama 3.2 11B Vision Instruct at Q4_K_M quantization (7.8 GB), estimated ~123 tokens/sec.
← Can I Run It? | Llama 3.2 Family Model Page | NVIDIA GeForce RTX 5080 GPU Page | Check Your Hardware