Best Local LLMs for the NVIDIA GeForce RTX 3080 (10GB)

Written by Jakub Rusinowski · Last updated July 12, 2026

Best all-round pick: Qwen 3 8B

The NVIDIA GeForce RTX 3080 (10GB) has 10 GB of VRAM, of which about 10 GB is available to a model. The largest model it can hold is Qwen 3 14B (15B, 9.7 GB at Q4_K_M).

Best models overall on the NVIDIA GeForce RTX 3080 (10GB)

ModelScoreMemory at Q4_K_MContextLicence
Qwen 3 8B95.85.8 GB125KApache 2.0
Qwen 3.5 14B94.49.3 GB125KApache 2.0
Gemma 3 12B Instruct93.68 GB128KGemma
Qwen 3.5 14B93.59.3 GB125KApache 2.0
GLM-6 9B93.36.2 GB125KMIT
GLM-4.7 9B93.16.2 GB125KApache-2.0

Best model by what you are doing

WorkloadRecommended on this GPU
CodingQwen3-Coder 8B (95.9)
Qwen 3 14B (95.9)
DeepSeek R1 Distill Qwen 14B (94.2)
General assistantQwen 3 8B (95.8)
Qwen 3.5 14B (94.4)
Gemma 3 12B Instruct (93.6)
ReasoningQwen 3 14B (96)
DeepSeek R1 Distill Llama 8B (94.5)
Qwen 3.5 14B (94)
RAGQwen 3 14B (84.4)
Qwen 2.5 14B Instruct (81.8)
Qwen 3.5 14B (81.5)
AgentsGLM-6 9B (86.8)
GLM-5 9B (85.4)
Qwen 3.5 14B (84.6)
VisionQwen 3.5 14B (95.7)
Gemma 4 12B (92.3)
Gemma 4 12B (Unified) (92)

What the NVIDIA GeForce RTX 3080 (10GB) cannot run

These models are strong picks generally but exceed the 10 GB this card makes available.

ModelNeeds at Q4_K_MShort by
Gemma 4 27B ⭐17.2 GB~7.2 GB
Mistral Small 3.1 24B15.1 GB~5.1 GB
Qwen 3.7 35B-A3B22 GB~12 GB
Gemma 4 31B19.6 GB~9.6 GB

How these numbers are calculated

FAQ

What is the best LLM for the NVIDIA GeForce RTX 3080 (10GB)?

Qwen 3 8B is the strongest all-round pick that fits its 10 GB.

What is the largest model the NVIDIA GeForce RTX 3080 (10GB) can run?

Qwen 3 14B — 15B parameters, needing 9.7 GB at Q4_K_M.

How much of the NVIDIA GeForce RTX 3080 (10GB)'s memory can a model actually use?

About 10 GB of its 10 GB, before the desktop and runtime overhead are accounted for.

Compatibility Checks for the NVIDIA GeForce RTX 3080 (10GB)

Similar GPUs

By Workload

More

← All GPUs | NVIDIA GeForce RTX 3080 (10GB) specs