Nex-N1 — Nex-AGI 的本地 AI 模型
作者: Jakub Rusinowski · 最后更新:
Nex-AGI's first agentic series, and the one that set the pattern: take someone else's open base weights, post-train them for tool use and autonomy, publish the result. The flagship is a DeepSeek-V3.1 post-train scoring 80.2 on tau2-Bench and 70.6 on SWE-bench Verified. Superseded by Nex-N2, but kept because the question 'what does a 671B agentic post-train need?' has a factual answer that does not change when the model stops being recommended.
变体
Nex-N1 最小的变体在 Q4_K_M 下约需 406 GB 显存——量化权重加框架开销,不含 KV 缓存。
| 模型 | Q4 下显存 | 显存 | 上下文 | 运行 |
|---|---|---|---|---|
| DeepSeek-V3.1-Nex-N1 671B → 671B (37B active) | ~405.9 GB | 128,000 | ollama run nex-n1 |
显存为 Q4_K_M 下的量化权重加开销,与 GPU 与显存检测器使用同一引擎计算。
如何在本地运行 Nex-N1
安装 Ollama,然后拉取标签。
ollama run nex-n1在上方选择一个尺寸,查看它自己的显存、速度估算和安装命令。
许可证
Commercial use permitted. No usage restrictions beyond attribution.
适用于: DeepSeek-V3.1-Nex-N1 671B推荐 GPU
目录中能在本地运行 Nex-N1(至少 406 GB 显存)的最便宜 GPU 是 Apple M3 Ultra (512 GB).
我的 GPU 能运行 Nex-N1 吗?
Nex-N1 — 常见问题
How much VRAM does Nex-N1 need?
Nex-N1 needs about 406 GB VRAM at Q4_K_M quantization for its smallest variant. Variants: DeepSeek-V3.1-Nex-N1 671B (406 GB, Q4_K_M). On Apple Silicon, unified memory counts toward this requirement.
Can I run Nex-N1 on an RTX 4090 (24 GB)?
Nex-N1's smallest variant needs about 406 GB, which exceeds a single RTX 4090 (24 GB). Use multiple GPUs, a higher-VRAM card, or Apple Silicon with large unified memory.
What quantization should I use for Nex-N1?
Q4_K_M is the best balance of quality and VRAM for Nex-N1 in most cases. Choose Q8_0 for near-lossless quality if you have spare VRAM, or smaller quants (Q3/Q2) only when memory is tight.
How do I run Nex-N1 with Ollama?
Nex-N1 has no local Ollama tag — the published tag is cloud-hosted, so running it sends your prompts to a hosted GPU rather than your own machine.