Models you can run yourself

A model, every quantisation somebody made of it, and what each one costs to run.

Behind2 sourcesupdated 4h agofresh within 24h
3,231things
838named by two sources or more
33conflicts
2sources

Only one source knows: GGUF quantisations 2326 · Hugging Face models 67

Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
PropertyHugging Face modelsGGUF quantisations
LicencecardData.licensecardData.license

Everything else the sources say is shown side by side, and not compared.

kind=quantisation ✕

EverythingQuantised, and openly licensedmodel 1645quantisation 9343

Found

26–50 of 5,539
ThingKindLicenceDownloadsDate
bartowski/MiMo-V2.6-Distill-Qwen-9B-GGUF
XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B · GGUF quantisations
quantisation 36 apache-2.002026-09-21
nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2-GGUF
· GGUF quantisations
quantisation 1 apache-2.01026862026-09-03
audreyt/DeepSeek-V4.1-Flash-Abliterated-GGUF
s-zaizen/DeepSeek-V4.1-Flash-Abliterated · GGUF quantisations
quantisation 1 mit1019452026-09-12
OS-Software/Ternary-Bonsai-2-27B-Uncensored-Heretic-GGUF
prism-ml/Ternary-Bonsai-2-27B-gguf · GGUF quantisations
quantisation 37 apache-2.002026-09-17
peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF-MTP
huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated · GGUF quantisations
quantisation 5 apache-2.0153072026-08-30
llmfan46/Qwen3.8-27B-Ultra-Uncensored-Heretic-Native-MTP-Preserved-GGUF
llmfan46/Qwen3.8-27B-Ultra-Uncensored-Heretic-Native-MTP-Preserved · GGUF quantisations
quantisation 8 apache-2.01402026-08-30
ukisai/Swift-1.5-Qwen3.8-27B-GSQ-RCO-GGUF
ukisai/Swift-1.5-Qwen3.8-27b · GGUF quantisations
quantisation 13 other02026-09-21
DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored
DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored · Hugging Face models, GGUF quantisations
model 2 quantisation 7 apache-2.01081 / 1127 / 1407 / 1552 / 2002 / 4241 / 484 / 57515 / 6582026-09-09
audio-cpp/VibeVoice-ASR-Streaming-7B-GGUF
microsoft/VibeVoice-ASR-Streaming-7B · GGUF quantisations
quantisation 3 mit3182026-09-04
0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored
0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored · Hugging Face models, GGUF quantisations
model 2 quantisation 2 apache-2.0 / other1021 / 12829 / 55865 / 5862026-08-31
nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2
· GGUF quantisations
quantisation 1 apache-2.0550292026-09-03
WalkingCat/Soprano-1.1-80M-GGUF
· GGUF quantisations
quantisation 1 apache-2.0540762026-08-27
llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved-GGUF
llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved · GGUF quantisations
quantisation 1 mit535182026-08-29
abenzerps/Nex-N2.5-mini-GGUF
nex-agi/Nex-N2.5-mini · GGUF quantisations
quantisation 16 apache-2.002026-09-08
SC117/Qwen3.8-Flash-Next-GSQ-RCO-abliterated-GGUF
ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF · GGUF quantisations
quantisation 7 apache-2.002026-09-17
QuantaPlanta/MiMo-V2.6-Flash-MOPD-MXFP4-GGUF
XiaomiMiMo/MiMo-V2.6-Flash-MOPD · GGUF quantisations
quantisation 9 mit10812026-09-27
NANI-Nithin/K2-Horizon-MoVA-36B-A4B-GGUF
IFM/K2-Horizon-MoVA-36B-A4B · GGUF quantisations
quantisation 12 apache-2.010912026-09-03
dgpl/dgpl-experimental-1
· GGUF quantisations
quantisation 1 apache-2.0439582026-09-30
elichen-skymizer/GAP-models
· GGUF quantisations
quantisation 1 —434712026-09-07
prithivMLmods/Qwen-Image-2.1-PE-I2I-GGUF
Qwen/Qwen-Image-2.1-PE-I2I · GGUF quantisations
quantisation 3 other321642026-09-20
Jommarn/UNSEEN_Gemma_4_26B_NSFW-GGUF
· GGUF quantisations
quantisation 1 —406132026-08-28
DuoNeural/Cyber-Ornith-1.5-9B-OBLITERATED
DuoNeural/Cyber-Ornith-1.5-9B-OBLITERATED · Hugging Face models, GGUF quantisations
model 2 quantisation 3 apache-2.00 / 1213 / 1342 / 1349 / 383912026-09-25
dealignai/GLM-5.3-Flash-UNCENSORED-FP8
dealignai/GLM-5.3-Flash-UNCENSORED-FP8 · Hugging Face models, GGUF quantisations
model 2 quantisation 6 mit1148 / 11978 / 1492 / 38203 / 47456 / 48688 / 5994 / 74262026-08-26
inclusionAI/Ling-3.0-tiny-GGUF
· GGUF quantisations
quantisation 1 mit373042026-08-30
slashreboot/athena-class-model-a
google/gemma-4-31B-it · GGUF quantisations
quantisation 19 apache-2.002026-09-04

Facets

Kind 10988 of 10988 claims

model1645

Source 10988 of 10988 claims

Licence 8499 of 10988 claims

mit744

Sources

Hugging Face models primary

The model itself: its licence, its task and how many people fetch it.

model · failing1645

GGUF quantisations high

Which quantisations exist, which is what decides whether it runs on the machine somebody has.

quantisation · current9343