Sized at Q4_K_M with a 8K context: weights, KV cache and framework overhead, against what Mac mini M4, 16GB leaves free. Gemma 3 12B Instruct at Q4_K_M needs 11.1 GB and you have 12 GB available, so it fits with 0.9 GB to spare.
Gemma 3 12B Instruct at Q4_K_M needs 11.1 GB and you have 12 GB available, so it fits with 0.9 GB to spare.
Your machine is not exactly this one. Run this for your exact setup — the form opens pre-filled with Mac mini M4, 16GB and Gemma 3 12B Instruct.
Yes. Gemma 3 12B Instruct at Q4_K_M needs 11.1 GB and you have 12 GB available, so it fits with 0.9 GB to spare.
Yes, and it is the figure people forget. The 11.1 GB above already includes the KV cache at 8K; that cache grows roughly linearly with context, so doubling the window adds real gigabytes rather than a rounding error. If you plan to work with long documents on Mac mini M4, 16GB, size for the context you will actually use, not the default.
Data behind this page last checked 2025-03-12.