Models you can run yourself

A model, every quantisation somebody made of it, and what each one costs to run.

Behind2 sourcesupdated 5h agofresh within 24h
3,060things
722named by two sources or more
28conflicts
2sources

Only one source knows: GGUF quantisations 2276 · Hugging Face models 62

Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
PropertyHugging Face modelsGGUF quantisations
LicencecardData.licensecardData.license

Everything else the sources say is shown side by side, and not compared.

source=zetlyn/models-gguf ✕kind=quantisation ✕

EverythingQuantised, and openly licensedmodel 1568quantisation 8921

Found

1–25 of 5,286
ThingKindLicenceDownloadsDate
prism-ml/Ternary-Bonsai-2-27B-gguf
Qwen/Qwen3.8-27B · GGUF quantisations
quantisation 273 apache-2.002026-08-26
DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU
DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU · Hugging Face models, GGUF quantisations
model 2 quantisation 53 apache-2.0 / cc-by-4.0 / other0 / 1015 / 1016 / 10199 / 1088 / 12341 / 12717 / 1379 / 14563 / 1476 / 1511 / 1557 / 1559 / 15592 / 1563 / 159 / 1634 / 1683597 / 1719 / 172 / 1724 / 1742 / 1754 / 1787 / 1790 / 17942 / 1840 / 1894 / 1910 / 2027 / 218 / 2217 / 2268 / 2311 / 2572 / 2638 / 29374 / 3160 / 31661 / 3255 / 3367 / 3375 / 3470 / 3592 / 438 / 4756 / 506 / 53 / 5401 / 5748 / 6578 / 703 / 8504 / 904 / 9102026-08-25
deepseek-ai/DeepSeek-V4.1-Flash
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face models, GGUF quantisations
model 2 quantisation 17 apache-2.0 / mit0 / 1099587 / 12061 / 158 / 15890 / 1693 / 1737 / 19701 / 2157 / 25057 / 257 / 3550 / 41756 / 5384 / 640577 / 767871 / 86 / 96852026-09-10
abenzerps/Qwen-Image-2.1-Uncensored-GGUF
Qwen/Qwen-Image-2.1 · GGUF quantisations
quantisation 68 apache-2.002026-09-20
Serveurperso/YuE2-GGUF
m-a-p/YuE2-3B · GGUF quantisations
quantisation 9 cc-by-nc-4.01280782026-09-10
Serveurperso/YuE2-GGUF
m-a-p/YuE2-Vae · GGUF quantisations
quantisation 3 cc-by-nc-4.06562026-09-12
Serveurperso/YuE2-GGUF
m-a-p/SheetSage2 · GGUF quantisations
quantisation 4 cc-by-nc-4.0173442026-09-12
Serveurperso/YuE2-GGUF
m-a-p/MERT-v2-FullSong · GGUF quantisations
quantisation 2 cc-by-nc-4.07510042026-09-12
unsloth/GLM-5.3-GGUF
zai-org/GLM-5.3 · GGUF quantisations
quantisation 13 mit02026-08-28
CMSManhattan/JiRackDeltaNet_27b
· GGUF quantisations
quantisation 1 mit5520252026-08-29
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF
Qwen/Qwen3.8-Flash-Next · GGUF quantisations
quantisation 126 apache-2.002026-08-26
OBLITERATUS/Ornith-1.5-9B-OBLITERATED
ornith-ai/Ornith-1.5-9B · GGUF quantisations
quantisation 24 apache-2.002026-08-27
XHToken/Spark-X2.5-4B-GGUF
XHToken/Spark-X2.5-4B · GGUF quantisations
quantisation 33 apache-2.002026-08-28
AngelSlim/Hy4-preview-GGUF
· GGUF quantisations
quantisation 1 —3569562026-08-28
empero-ai/Qwen3.8-35B-A3B-Distill
empero-ai/Qwen3.8-35B-A3B-Distill · Hugging Face models, GGUF quantisations
model 2 quantisation 11 apache-2.01445 / 1756 / 19162 / 21029 / 290 / 303921 / 436 / 4825 / 513 / 5362 / 5468 / 5691 / 99292026-09-16
openbmb/MiniCPM5-2B-GGUF
· GGUF quantisations
quantisation 1 apache-2.02346132026-09-05
ukisai/Swift-Qwen3.8-27B-GGUF
ukisai/Swift-Qwen3.8-27b · GGUF quantisations
quantisation 18 apache-2.002026-09-11
DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored
DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored · Hugging Face models, GGUF quantisations
model 2 quantisation 11 apache-2.015792 / 1821 / 1917 / 1970 / 203773 / 2165 / 2332 / 2670 / 3706 / 5187 / 613 / 6902 / 9692026-08-31
antirez/glm-5.3-flash-gguf
· GGUF quantisations
quantisation 1 —1922432026-08-27
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp · Hugging Face models, GGUF quantisations
model 2 quantisation 16 apache-2.0 / mit / other11158 / 13551 / 189004 / 212 / 2193 / 2498 / 2559 / 3355 / 36695 / 442 / 5618 / 588 / 62318 / 813 / 925431 / 932 / 951084 / 96012026-08-31
XHToken/Spark-X2.5-1.7B-GGUF
XHToken/Spark-X2.5-1.7B · GGUF quantisations
quantisation 13 apache-2.012022026-08-28
pottokao/Qwen-Image-2.1-Text-Encoder-Heretic-GGUF
pottokao/Qwen-Image-2.1-Text-Encoder-Heretic · GGUF quantisations
quantisation 6 apache-2.013262026-09-20
Accio-Lab/occamy-1.0
Accio-Lab/occamy-1.0 · Hugging Face models, GGUF quantisations
model 2 quantisation 13 apache-2.01010 / 1025 / 157 / 158222 / 1759 / 2440 / 2564 / 3431 / 4511 / 4652 / 6035 / 6115 / 8117 / 862 / 91352026-08-13
bartowski/orcarouter_Qwen3.8-27B-Uncensored-GGUF
orcarouter/Qwen3.8-27B-Uncensored · GGUF quantisations
quantisation 15 apache-2.01495352026-08-27
mradermacher/Qwen3.8-Flash-Next-Uncensored-GGUF
orcarouter/Qwen3.8-Flash-Next-Uncensored · GGUF quantisations
quantisation 23 apache-2.002026-08-27
← PreviousPage 1 of 212Next →

Facets

Kind 10489 of 10489 claims

model1568

Source 10489 of 10489 claims

Licence 8092 of 10489 claims

mit713

Sources

Hugging Face models primary

The model itself: its licence, its task and how many people fetch it.

model · failing1568

GGUF quantisations high

Which quantisations exist, which is what decides whether it runs on the machine somebody has.

quantisation · current8921