HomeModels

Model sheets

Every local model, measured

Every model sheet is built from measurements, not from its marketing page: the real size of the published files, the context memory computed from the attention pattern it declares layer by layer, and what it costs in tokens to say the same thing in four languages.

The catalogue Measured
Models with a sheet18
Smallest / largest1.03 / 433.83 GiB
Fit in 8 GB5 of 18
Costliest languageDE ×1.62
Measured on2026-08-25
ModelSizeCompression shownGerman costFits in 8 GBLicense
deepseek-v4-flash127.28 GiBIQ4_XS×1.58nomit
gemma4-12b6.63 GiBQ4_K_M×1.33noapache-2.0
gemma4-26b-a4b15.78 GiBQ4_K_M×1.33noapache-2.0
gemma4-31b17.07 GiBQ4_K_M×1.33noapache-2.0
gemma4-e2b2.89 GiBQ4_K_M×1.33yesapache-2.0
gemma4-e4b4.64 GiBQ4_K_M×1.33yesapache-2.0
glm-5.2433.83 GiBQ4_K_M×1.49nomit
gpt-oss-120b58.46 GiBQ4_K_M×1.29noapache-2.0
gpt-oss-20b10.83 GiBQ4_K_M×1.29noapache-2.0
llama4-scout-17b60.87 GiBQ4_K_M×1.30noother
mistral-small-24b13.35 GiBQ4_K_M×1.40noapache-2.0
qwen3-1.7b1.03 GiBQ4_K_M×1.62yesapache-2.0
qwen3-4b-25072.33 GiBQ4_K_M×1.62yesapache-2.0
qwen3-8b4.68 GiBQ4_K_M×1.62yesapache-2.0
qwen3-coder-30b17.28 GiBQ4_K_M×1.62noapache-2.0
qwen3.6-27b15.66 GiBQ4_K_M×1.31noapache-2.0
qwen3.6-35b-a3b20.61 GiBQ4_K_M×1.31noapache-2.0
qwen3.8-27b15.33 GiBQ4_K_M×1.31noapache-2.0

Sizes are the real published files at Q4_K_M, our quality floor. Where the author does not publish it, the row falls back to the nearest compression and says which one — a size at one compression compared against another is not a comparison.

This table does not say what runs on your machine. It says how big each model is. Which of them fit in your memory — with the context your work needs, and in your language — is what the Model Finder answers.

Check it against my machine