作者: Jakub Rusinowski · 最后更新: 2026年7月30日
The complete guide to Alibaba's Qwen: every model size explained across Qwen 3 and 3.5, thinking mode, MoE and Gated DeltaNet, native vision, 201 languages, which one fits your GPU, and how it compares to DeepSeek, Gemma and Llama.
If you run open models locally, you will eventually run Qwen. Alibaba's family has quietly become the default recommendation for anyone with a consumer GPU — not because it wins one headline benchmark, but because it is the most complete open family available: sizes from a phone-friendly 0.8B up to frontier-class hundreds of billions, native vision, 201 languages, a genuine reasoning mode, one of the most permissive licenses in AI, and — crucially — a size that actually fits and runs fast on the hardware you own. This is the complete guide to Qwen. We will trace the family from Qwen 2.5 throug…