Models you can run yourself
A model, every quantisation somebody made of it, and what each one costs to run.
Only one source knows: GGUF quantisations 2276 · Hugging Face models 62
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
| Property | Hugging Face models | GGUF quantisations |
|---|---|---|
| Licence | cardData.license | cardData.license |
Everything else the sources say is shown side by side, and not compared.
Found
1–25 of 107| Thing | Kind | Licence | Downloads | Date |
|---|---|---|---|---|
| CMSManhattan/JiRackDeltaNet_27b · GGUF quantisations | quantisation 1 | mit | 552025 | 2026-08-29 |
| OBLITERATUS/Ornith-1.5-9B-OBLITERATED ornith-ai/Ornith-1.5-9B · GGUF quantisations | quantisation 18 | apache-2.0 | 1005 | 2026-08-27 |
| peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF-MTP huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated · GGUF quantisations | quantisation 3 | mit | 15307 | 2026-08-31 |
| deepseek-ai/DeepSeek-V4-Flash-Vision-Exp deepseek-ai/DeepSeek-V4-Flash-Vision-Exp · Hugging Face models, GGUF quantisations | model 2 quantisation 16 | apache-2.0 / mit / other | 11158 / 13551 / 189004 / 212 / 2193 / 2498 / 2559 / 3355 / 36695 / 442 / 5618 / 588 / 62318 / 813 / 925431 / 932 / 951084 / 9601 | 2026-08-31 |
| llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved-GGUF llmfan46/LongCat-Flash-Lite-Sparse-Ultra-Uncensored-Heretic-Native-MTP-And-LSA-Preserved · GGUF quantisations | quantisation 1 | mit | 53518 | 2026-08-29 |
| inclusionAI/Ling-3.0-tiny-GGUF · GGUF quantisations | quantisation 1 | mit | 37304 | 2026-08-30 |
| orcarouter/GLM-5.3-Flash-Uncensored-GGUF zai-org/GLM-5.3-Flash · GGUF quantisations | quantisation 45 | mit | 0 | 2026-08-27 |
| audio-cpp/VibeVoice-7B-GGUF vibevoice/VibeVoice-7B · GGUF quantisations | quantisation 1 | mit | 28476 | 2026-08-28 |
| BoldingBuilds/orcarouter_GLM-5.3-Flash-Uncensored-GGUF orcarouter/GLM-5.3-Flash-Uncensored-FP8 · GGUF quantisations | quantisation 2 | mit | 23894 | 2026-08-31 |
| LuffyTheFox/Tiel-Coder-35B-A3B-Genesis-Hermes-GGUF ornith-ai/Ornith-1.5-35B-A3B · GGUF quantisations | quantisation 27 | apache-2.0 | 0 | 2026-08-27 |
| mradermacher/Ornith-1.5-35B-A3B-FULLY-OBLITERATED-GGUF jhone888/Ornith-1.5-35B-A3B-FULLY-OBLITERATED · GGUF quantisations | quantisation 2 | mit | 10021 | 2026-08-30 |
| meshllm/GLM-5.3-Flash-UD-Q4_K_XL-layers unsloth/GLM-5.3-Flash-GGUF · GGUF quantisations | quantisation 4 | mit | 1119 | 2026-08-28 |
| inclusionAI/Ling-3.0-flash-GGUF · GGUF quantisations | quantisation 1 | mit | 8979 | 2026-08-31 |
| deepseek-ai/DeepSeek-V4-Flash-0731 deepseek-ai/DeepSeek-V4-Flash-0731 · Hugging Face models, GGUF quantisations | model 2 quantisation 5 | mit / other | 3824182 / 4 / 4494615 / 5862 / 591 / 6392 / 855 | 2026-07-31 |
| llmfan46/LongCat-Flash-Lite-Sparse-Uncensored-Heretic-Native-MTP-And-LSA-Preserved-GGUF llmfan46/LongCat-Flash-Lite-Sparse-Uncensored-Heretic-Native-MTP-And-LSA-Preserved · GGUF quantisations | quantisation 1 | mit | 5317 | 2026-08-29 |
| Tinker-Stack/DeepSeek-V4-Flash-0731-REAP-150b-TQ3_4S-GGUF puwaer/DeepSeek-V4-Flash-0731-reap-150b · GGUF quantisations | quantisation 1 | mit | 5269 | 2026-08-29 |
| 3MPER0RR/Ornith1.5-9B-3MPER0RR-abliterated 3MPER0RR/Ornith1.5-9B-3MPER0RR-abliterated · Hugging Face models, GGUF quantisations | model 2 quantisation 4 | mit | 1522 / 278 / 3501 / 4283 / 432 / 831 | 2026-08-31 |
| mradermacher/Aetheris-Core-2B-i1-GGUF Miiyamoto255/Aetheris-Core-2B · GGUF quantisations | quantisation 2 | mit | 3190 | 2026-08-31 |
| mradermacher/Ornith-1.5-9B-OBLITERATED-i1-GGUF OBLITERATUS/Ornith-1.5-9B-OBLITERATED · GGUF quantisations | quantisation 3 | apache-2.0 | 1278 | 2026-08-27 |
| orcarouter/DeepSeek-V4-Flash-Vision-Uncensored-GGUF orcarouter/DeepSeek-V4-Flash-Vision-Uncensored · GGUF quantisations | quantisation 2 | mit | 1363 | 2026-09-02 |
| ChonkE/Granite_42_30b_Abliterated ChonkE/Granite_42_30b_Abliterated · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | mit | 2093 / 25 / 42 / 426 | 2026-08-28 |
| mradermacher/InfoDensity-DeepSeek-R1-Distill-Qwen-1.5B-i1-GGUF amao0o0/InfoDensity-DeepSeek-R1-Distill-Qwen-1.5B · GGUF quantisations | quantisation 2 | mit | 1982 | 2026-09-02 |
| speleoalex/physisml-it-preview · GGUF quantisations | quantisation 1 | mit | 1737 | 2026-08-28 |
| dots-studio/dots.mocr dots-studio/dots.mocr · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | mit | 120 / 1499 / 384783 / 423053 | 2026-03-19 |
| rumisan/DeepSeek-R1-Distill-Qwen-1.5B-Fully-Uncensored-i1-GGUF nicoboss/DeepSeek-R1-Distill-Qwen-1.5B-Fully-Uncensored · GGUF quantisations | quantisation 2 | mit | 1221 | 2026-08-29 |
Facets
Licence 8092 of 10489 claims
mit895
Sources
Hugging Face models primary
The model itself: its licence, its task and how many people fetch it.
model · failing1568
GGUF quantisations high
Which quantisations exist, which is what decides whether it runs on the machine somebody has.
quantisation · current8921