作者: Jakub Rusinowski · 最后更新: 2026年3月1日
DeepSeek 2026年初的V3进化版。V3.2(671B总参数,37B活跃参数)专注于改进长上下文推理、工具使用和智能体任务。在大多数基准测试中对标GPT-4o级别性能,MIT许可证完全开源。
| Licence | What it permits | Applies to |
|---|---|---|
MIT | Commercial use permitted Commercial use permitted. No usage restrictions beyond attribution. | DeepSeek V3.2 671B, DeepSeek V3.2 685B |
| DeepSeek V3.2 671B | Min 406 GB VRAM · Q4_K_M · 128,000 ctx · ollama run deepseek-v3.2:671b-q4 |
| DeepSeek V3.2 685B | Min 414 GB VRAM · Q4_K_M · 128,000 ctx · |
The cheapest GPU that runs DeepSeek V3.2 locally (min 406 GB VRAM) is the Apple M3 Ultra (512 GB).
Install Ollama then run: ollama run deepseek-v3.2:671b-q4
Minimum VRAM: 406 GB. For best results use Q4_K_M quantization.
DeepSeek V3.2 needs about 406 GB VRAM at Q4_K_M quantization for its smallest variant. Variants: DeepSeek V3.2 671B (406 GB, Q4_K_M); DeepSeek V3.2 685B (414 GB, Q4_K_M). On Apple Silicon, unified memory counts toward this requirement.
DeepSeek V3.2's smallest variant needs about 406 GB, which exceeds a single RTX 4090 (24 GB). Use multiple GPUs, a higher-VRAM card, or Apple Silicon with large unified memory.
Q4_K_M is the best balance of quality and VRAM for DeepSeek V3.2 in most cases. Choose Q8_0 for near-lossless quality if you have spare VRAM, or smaller quants (Q3/Q2) only when memory is tight.
Install Ollama, then run: ollama run deepseek-v3.2:671b-q4. This downloads DeepSeek V3.2 and starts a local, OpenAI-compatible endpoint — no internet connection is needed after the initial download.