Will it run?

Pick a model and a card. You get every shipped quantization against every context length and KV cache dtype — each cell computed independently — rather than one yes-or-no that hides the settings it assumed.

Popular models

then choose a card on the model page

Or start from your hardware