Models you can run yourself

A model, every quantisation somebody made of it, and what each one costs to run.

Behind2 sourcesthe promise of 24h does not hold
3,085things
793named by two sources or more
32conflicts
2sources

Only one source knows: GGUF quantisations 2229 · Hugging Face models 63

Behind: zetlyn/models-hf has not completed an update.
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
PropertyHugging Face modelsGGUF quantisations
LicencecardData.licensecardData.license

Everything else the sources say is shown side by side, and not compared.

kind=quantisation ✕licence=mit ✕licence=mit ✕kind=quantisation ✕

EverythingQuantised, and openly licensedmodel 1568quantisation 9011

Found

1–25 of 469
ThingKindLicenceDownloadsDate
deepseek-ai/DeepSeek-V4.1-Flash
deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face models, GGUF quantisations
model 2 quantisation 17 apache-2.0 / mit0 / 1099587 / 12061 / 158 / 15890 / 1693 / 1737 / 19701 / 2157 / 25057 / 257 / 3550 / 41756 / 5384 / 640577 / 767871 / 86 / 96852026-09-10
CMSManhattan/JiRackDeltaNet_27b
· GGUF quantisations
quantisation 1 mit5520252026-08-29
OBLITERATUS/Ornith-1.5-9B-OBLITERATED
ornith-ai/Ornith-1.5-9B · GGUF quantisations
quantisation 20 apache-2.010052026-08-27
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp · Hugging Face models, GGUF quantisations
model 2 quantisation 16 apache-2.0 / mit / other11158 / 13551 / 189004 / 212 / 2193 / 2498 / 2559 / 3355 / 36695 / 442 / 5618 / 588 / 62318 / 813 / 925431 / 932 / 951084 / 96012026-08-31
audreyt/DeepSeek-V4.1-Flash-Abliterated-GGUF
s-zaizen/DeepSeek-V4.1-Flash-Abliterated · GGUF quantisations
quantisation 1 mit1019452026-09-12
peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF-MTP
huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated · GGUF quantisations
quantisation 3 mit153072026-08-31
audio-cpp/VibeVoice-ASR-Streaming-7B-GGUF
microsoft/VibeVoice-ASR-Streaming-7B · GGUF quantisations
quantisation 3 mit3182026-09-04
llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved-GGUF
llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved · GGUF quantisations
quantisation 1 mit535182026-08-29
QuantaPlanta/MiMo-V2.6-Flash-MOPD-MXFP4-GGUF
XiaomiMiMo/MiMo-V2.6-Flash-MOPD · GGUF quantisations
quantisation 9 mit1612026-09-27
inclusionAI/Ling-3.0-tiny-GGUF
· GGUF quantisations
quantisation 1 mit373042026-08-30
orcarouter/GLM-5.3-Flash-Uncensored-GGUF
zai-org/GLM-5.3-Flash · GGUF quantisations
quantisation 46 mit02026-08-27
audio-cpp/AuK-Base-and-Flash-GGUF
tencent/AuK · GGUF quantisations
quantisation 1 mit315192026-09-20
dealignai/DeepSeek-V4.1-Flash-UNCENSORED-FP8
dealignai/DeepSeek-V4.1-Flash-UNCENSORED-FP8 · Hugging Face models, GGUF quantisations
model 2 quantisation 1 mit30063 / 39227 / 559902026-09-10
audio-cpp/VibeVoice-7B-GGUF
vibevoice/VibeVoice-7B · GGUF quantisations
quantisation 1 mit284762026-08-28
DogContext/GLM-5.3-Flash-Uncensored-Q2-ds4
orcarouter/GLM-5.3-Flash-Uncensored-FP8 · GGUF quantisations
quantisation 3 mit11482026-08-31
ggml-org/MiMo-V2.6-Distill-Qwen-9B-GGUF
XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B · GGUF quantisations
quantisation 34 apache-2.002026-09-21
gbuzhf/Ornith-1.5-35B-A3B-Abliterated-CyberTiel-Calibrated-MTPv2-ICE-GGUF
peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF-MTP · GGUF quantisations
quantisation 3 mit153072026-09-10
LuffyTheFox/Tiel-Coder-35B-A3B-Genesis-Hermes-GGUF
ornith-ai/Ornith-1.5-35B-A3B · GGUF quantisations
quantisation 29 apache-2.002026-08-27
christopherthompson81/VibeVoice-ASR-Streaming-1.5B-GGUF
microsoft/VibeVoice-ASR-Streaming-1.5B · GGUF quantisations
quantisation 1 mit127652026-09-19
bartowski/Ling-3.0-flash-Fin-GGUF
inclusionAI/Ling-3.0-flash-Fin · GGUF quantisations
quantisation 3 mit125662026-09-03
dealignai/GLM-5.3-Flash-UNCENSORED-FP8
dealignai/GLM-5.3-Flash-UNCENSORED-FP8 · Hugging Face models, GGUF quantisations
model 2 quantisation 6 mit1148 / 11978 / 1492 / 38203 / 47360 / 47456 / 5994 / 74262026-08-26
ggml-org/MiMo-V2.6-Flash-RL-GGUF
XiaomiMiMo/MiMo-V2.6-Flash-RL · GGUF quantisations
quantisation 15 mit02026-09-22
mradermacher/Ornith-1.5-35B-A3B-FULLY-OBLITERATED-GGUF
jhone888/Ornith-1.5-35B-A3B-FULLY-OBLITERATED · GGUF quantisations
quantisation 2 mit100212026-08-30
meshllm/GLM-5.3-Flash-UD-Q4_K_XL-layers
unsloth/GLM-5.3-Flash-GGUF · GGUF quantisations
quantisation 6 mit10812026-08-28
bloomer010/Ling-3.0-flash-VL-GGUF
inclusionAI/Ling-3.0-flash-VL · GGUF quantisations
quantisation 6 mit02026-09-07
← PreviousPage 1 of 19Next →

Facets

Kind 10579 of 10579 claims

model1568

Source 10579 of 10579 claims

Licence 8160 of 10579 claims

mit718

Sources

Hugging Face models primary

The model itself: its licence, its task and how many people fetch it.

model · partial1568

GGUF quantisations high

Which quantisations exist, which is what decides whether it runs on the machine somebody has.

quantisation · current9011