Apple · apple
Apple M2
Apple M2 has 24 GB of unified memory at 102 GB/s — about 16.74 GiB usable after driver and compositor overhead. 1794 of 2118 indexed models fit at 64K context with q8_0 KV. Note only 18 GB of its 24 GB is allocatable to the GPU.
Spec sheet· bandwidth, theoreticalFrom the file· fit from summed bytesPredicted· speed
Memory
24 GB
LPDDR5-6400
Bandwidth
102 GB/s
128-bit bus
Tensor FP16
—
dense
TDP
—
text 1521vision language 170video 16audio asr 39image 1embedding 26audio tts 21
What fits at 64K context
largest quantization that fits, per model · 1794 of 2118 indexed
| Model | Best quant | Params | Weights● | KV● | Total in memory◐ | Headroom◐ | tok/s◐ |
|---|---|---|---|---|---|---|---|
| dolphin-2.6-mixtral-8x7bMoE | I1-IQ2_S | 46.7B | 13.16 GiB | 4.25 GiB | 17.99 GiB | 0.01 GiB | 6±37% |
| xLAM-8x7b-rMoE | IQ2_S | 46.7B | 13.16 GiB | 4.25 GiB | 17.99 GiB | 0.01 GiB | 6±37% |
| Phi-3-medium-4k-instruct | Q6_K_L | 14.0B | 10.74 GiB | 6.64 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Ling-mini-2.0MoE | Q8_0 | 16.3B | 16.12 GiB | 1.33 GiB | 17.99 GiB | 0.01 GiB | 16±37% |
| Transformed-Journey-24B | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Mergedonia-AETHER-24B-v1a | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Mergedonia-AETHER-24B-v1b | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Slimaki-Tavern-24B-v1.3 | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Maginum-Cydoms-24B | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Maginum-Cydoms-24B-absolute-heresy | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Mistral-Small-3.2-24B-Instruct-2506-ultra-uncensored-heretic | IQ4_XS | 24.0B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Huihui-Mistral-Small-3.2-24B-Instruct-2506-abliterated-llamacppfixed | IQ4_XS | 24.0B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Mistral-Small-3_2-24B-Instruct-2506-antislop.v2 | IQ4_XS | 24.0B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Goetia-24B-v1.1 | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| MagiSeek-Pro-V1 | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Cogidonia-v2-24B | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Precog-24B-v1 | IQ4_XS | — | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| experiment024b | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Cydonia-24B-v4.3-absolute-heresy | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Cydonia-24B-v4.3-heretic-v2 | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Cydonia-24B-v4.3-heretic | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Cydonia-24B-v4.3-heretic-v4 | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Journeys-End-24B | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Dolphin-Mistral-GLM-4.7-Flash-24B-Venice-Edition-Thinking-Uncensored | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| WeirdCompound-v1.7-24b | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| Mistral-Small-24B-Instruct-Jbliterated | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| grok-oss-Apollyon-24B | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| grok-oss-Apollyon-24B-heretic | IQ4_XS | 23.6B | 12.00 GiB | 5.31 GiB | 17.99 GiB | 0.01 GiB | 5±8.3% |
| gemma-4-31B-it-qat-q4_0-unquantized-heretic | Q2_K_L | 31.3B | 11.42 GiB | 5.94 GiB | 17.98 GiB | 0.02 GiB | 5±8.3% |
| gemma-2-27b-it | IQ3_XS | 27.2B | 10.76 GiB | 6.54 GiB | 17.98 GiB | 0.02 GiB | 5±8.3% |
| magnum-v4-27b | IQ3_XS | 27.2B | 10.76 GiB | 6.54 GiB | 17.98 GiB | 0.02 GiB | 5±8.3% |
| Wan2.1-VACE-14B | Q8_0 | 17.3B | 17.38 GiB | 0.00 GiB | 17.97 GiB | 0.03 GiB | 5±8.3% |
| GRM-2.6-Plus-0628 | Q4_0 | 27.8B | 15.23 GiB | 2.13 GiB | 17.96 GiB | 0.04 GiB | 5±8.3% |
| ThinkingCap-Qwen3.6-27B | Q4_0 | 27.4B | 15.23 GiB | 2.13 GiB | 17.96 GiB | 0.04 GiB | 5±8.3% |
| Tess-4-27B | Q4_0 | 27.8B | 15.23 GiB | 2.13 GiB | 17.96 GiB | 0.04 GiB | 5±8.3% |
| Qwen3.8-27B | IQ4_NL | 27.8B | 15.22 GiB | 2.13 GiB | 17.95 GiB | 0.05 GiB | 5±8.3% |
| Qwen3.6-27B | IQ4_NL | 27.8B | 15.22 GiB | 2.13 GiB | 17.95 GiB | 0.05 GiB | 5±8.3% |
| InternVL3_5-30B-A3B | Q4_K_M | 30.8B | 17.35 GiB | 0.00 GiB | 17.95 GiB | 0.05 GiB | 5±8.3% |
| Skyfall-31B-v4.2 | IQ2_S | 31.4B | 10.10 GiB | 7.17 GiB | 17.94 GiB | 0.06 GiB | 5±8.3% |
| UncensoredLM-DeepSeek-R1-Distill-Qwen-14B | Q6_K_L | 14.2B | 11.22 GiB | 6.11 GiB | 17.93 GiB | 0.07 GiB | 5±8.3% |
| ALIA-40b-fc-2606 | I1-IQ2_XXS | 40.4B | 10.89 GiB | 6.38 GiB | 17.93 GiB | 0.07 GiB | 5±8.3% |
| ALIA-40b-instruct-2606 | I1-IQ2_XXS | 40.4B | 10.89 GiB | 6.38 GiB | 17.93 GiB | 0.07 GiB | 5±8.3% |
| dolphin-2.9.2-Phi-3-MediumKV unresolved | Q6_K | 14.0B | 10.67 GiB | 6.64 GiB | 17.92 GiB | 0.08 GiB | 5±8.3% |
| Phi-3-medium-128k-instruct | Q6_K | 14.0B | 10.67 GiB | 6.64 GiB | 17.92 GiB | 0.08 GiB | 5±8.3% |
| gemma-7b | I1-IQ2_XXS | 8.5B | 2.41 GiB | 14.88 GiB | 17.91 GiB | 0.09 GiB | 5±8.3% |
| c4ai-command-r-08-2024 | Q2_K | 32.3B | 11.93 GiB | 5.31 GiB | 17.91 GiB | 0.09 GiB | 5±8.3% |
| Llama3.2-30B-A3B-II-Dark-Champion-INSTRUCT-Heretic-Abliterated-UncensoredMoE | I1-IQ3_XXS | 30.0B | 11.09 GiB | 6.24 GiB | 17.90 GiB | 0.10 GiB | 6±37% |
| Qwen3-VL-8B-Instruct-Heretic | I1-Q6_K | 8.8B | 12.53 GiB | 4.78 GiB | 17.89 GiB | 0.11 GiB | 5±8.3% |
| MiniCPM-V-4_5 | Q6_K | 8.7B | 12.53 GiB | 4.78 GiB | 17.89 GiB | 0.11 GiB | 5±8.3% |
| gemma-4-26B-A4B-itMoE | Q4_K_M | 26.5B | 15.87 GiB | 1.48 GiB | 17.89 GiB | 0.11 GiB | 5±8.3% |
| Devstral-Small-2-24B-Instruct-2512 | IQ4_XS | 24.0B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Mistral-Small-3.2-24B-Instruct-2506 | IQ4_XS | 24.0B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Devstral-Small-2507 | IQ4_XS | 23.6B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Devstral-Small-2505 | IQ4_XS | 23.6B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Magistral-Small-2509 | IQ4_XS | 24.0B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Magistral-Small-2507 | IQ4_XS | 23.6B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Mistral-Small-3.1-24B-Instruct-2503 | IQ4_XS | 24.0B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Magistral-Small-2506 | IQ4_XS | 23.6B | 11.90 GiB | 5.31 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
| Qwen3-Coder-REAP-25B-A3BMoE | Q4_K_M | 24.9B | 14.15 GiB | 3.19 GiB | 17.88 GiB | 0.12 GiB | 9±37% |
| OmniAtlas-Qwen3-30B-A3B | I1-Q4_K_M | 31.7B | 17.28 GiB | 0.00 GiB | 17.88 GiB | 0.12 GiB | 5±8.3% |
Speed is modeled, not measured: decode is memory-bandwidth bound, so tokens per second is bytes read per token against achievable bandwidth. Mixture-of-experts models carry a wider band because only the routed experts are read each step, and few have been measured publicly.
Questions people ask
- What AI models can a Apple M2 run?
- 1794 of 2118 indexed open-weight models fit a Apple M2 at 65,536 context with q8_0 KV cache, the largest being dolphin-2.6-mixtral-8x7b at I1-IQ2_S. That covers text, vision-language, image, video and speech models.
- How much usable memory does a Apple M2 actually have?
- Its nameplate is 24 GB, but about 16.74 GiB is available to a model once driver and compositor overhead is accounted for, and only 18 GB of the pool can be allocated to the GPU at all.
- Is a Apple M2 fast for local AI?
- Its memory bandwidth is 102 GB/s, and that figure — not teraflops — is what governs token generation speed. Capacity decides what you can run; bandwidth decides how fast it runs.