// 6 issues · dated & citable
Local AI Reports
Biweekly digest issues on the state of local LLMs and hardware — new models, GPU price moves, tooling updates. Dated, citable, and archived.
Get each issue by emailThe Local AI Digest arrives every two weeks. No spam, unsubscribe anytime.Subscribe to the digest →
Latest · ReportLocal AI Report #6 — Best Local Coding Models: Three Picks for Every Hardware TierThe best local coding models for 8–16 GB, 24–32 GB and 128 GB machines: three picks per tier, what they run on, and how far behind Claude they score.Read the issue →
- Entry level (8–16 GB): K2 Horizon 7B. A 5.6 GB download that scores 21 on the independent Artificial Analysis index — nearly double Qwen3.5-9B's 11. Runners-up: Qwen3.5-9B (the easiest to install) and K2 Horizon 3.7B for 8 GB.
- Mid-range (24–32 GB): Qwen3.8-27B. Index 34, an 18 GB download, and about 21.9 GB at 64k context. Runners-up: K2 Horizon MoVA 36B-A4B (25, on a 32 GB card) and the dedicated coder Devstral Small 2.
- High-end (128 GB): Qwen3.8-Flash-Next. Index 40, a 120 GB download — tight on a DGX Spark or Strix Halo, too big for a 128 GB Mac. Runners-up: DeepSeek V4 Flash 0731 at 3-bit and Qwen3.8-27B at full precision.
- The gap to the subscriptions is real. The best model on one GPU scores 34; Claude Opus 5.5 scores 58 (index v4.3.2, checked 2026-10-08). That is the budget subscription tier, not the flagship.