Best local vision-language models

Models that accept images alongside text. Remember these need their vision projector loaded too, and that images consume context tokens quickly.

From the file· live filter over real dataFrom the file· 40 models

How this is ranked

Ranked by downloads among vision-language models with published quantizations.

Best local vision-language models

#ModelParamsSmallest quantSmallest quant
1Qwen3.6-35B-A3BMoE36.0B8.77 GiB8.77 GiB
2Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP27.8B10.12 GiB10.12 GiB
3Qwen3.5-9B9.7B2.97 GiB2.97 GiB
4gemma-4-26B-A4B-itMoE26.5B8.99 GiB8.99 GiB
5gemma-4-12B-it12.0B3.92 GiB3.92 GiB
6Qwen3.5-4B4.7B1.42 GiB1.42 GiB
7Qwythos-9B-Claude-Mythos-5-1M9.4B5.38 GiB5.38 GiB
8gemma-4-31B-it31.3B7.95 GiB7.95 GiB
9Muse-Glimmer-30B29.8B8.31 GiB8.31 GiB
10gemma-4-26B-A4B-it-qat-q4_0-unquantizedMoE26.5B13.45 GiB13.45 GiB
11Qwen3.5-122B-A10BMoE125B26.92 GiB26.92 GiB
12Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive34.7B34.37 GiB34.37 GiB
13Qwen3.5-0.8B873M0.31 GiB0.31 GiB
14gemma-4-E2B-it5.1B2.13 GiB2.13 GiB
15gemma-4-31B-it-qat-q4_0-unquantized32.7B16.44 GiB16.44 GiB
16gemma-4-E4B-it-qat-q4_0-unquantized7.9B4.80 GiB4.80 GiB
17Qwen3-VL-30B-A3B-InstructMoE31.1B7.60 GiB7.60 GiB
18Qwopus3.6-35B-A3B-v1MoE36.0B6.97 GiB6.97 GiB
19Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP9.7B9.21 GiB9.21 GiB
20Qwen3.5-35B-A3BMoE36.0B8.77 GiB8.77 GiB
21Qwythos-9B-v29.7B3.64 GiB3.64 GiB
22ThinkingCap-Qwen3.6-27B27.4B9.30 GiB9.30 GiB
23Qwopus3.6-27B-Coder27.8B8.89 GiB8.89 GiB
24Kimi-K2.7-CodeMoE1059B283.04 GiB283.04 GiB
25Qwen3-VL-4B-Instruct4.4B1.01 GiB1.01 GiB
26Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking39.5B13.76 GiB13.76 GiB
27Qwen2.5-VL-7B-Instruct8.3B1.93 GiB1.93 GiB
28Qwen3-VL-2B-Instruct2.1B0.50 GiB0.50 GiB
29Qwen3.5-27B27.8B7.98 GiB7.98 GiB
30gemma-3-12b-it12.2B2.85 GiB2.85 GiB
31Qwen3.5-2B2.3B0.72 GiB0.72 GiB
32InklingMoE952B210.69 GiB210.69 GiB
33Qwen3.6-27B-Heretic2-Uncensored-Finetune-Thinking27.4B9.77 GiB9.77 GiB
34Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-DistilledMoE36.0B12.34 GiB12.34 GiB
35Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-PreservedMoE35.1B16.08 GiB16.08 GiB
36Qwen3.5-9B9.7B4.31 GiB4.31 GiB
37Qwen3.6-35B-A3B-uncensored-hereticMoE35.1B15.71 GiB15.71 GiB
38gemma-4-31B-it-uncensored-heretic31.3B11.10 GiB11.10 GiB
39Step-3.7-Flash201B40.03 GiB40.03 GiB
40Qwopus3.6-27B-v227.8B12.39 GiB12.39 GiB
Spec sheetFrom the fileFrom the filewhat these mean