Models you can run yourself
A model, every quantisation somebody made of it, and what each one costs to run.
Only one source knows: GGUF quantisations 2276 · Hugging Face models 62
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
| Property | Hugging Face models | GGUF quantisations |
|---|---|---|
| Licence | cardData.license | cardData.license |
Everything else the sources say is shown side by side, and not compared.
Found
1–25 of 37| Thing | Kind | Licence | Downloads | Date |
|---|---|---|---|---|
| deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | other | 193 / 914582 | 2024-06-14 |
| deepseek-ai/DeepSeek-V2-Lite deepseek-ai/DeepSeek-V2-Lite · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | other | 0 / 223122 | 2024-05-15 |
| agentionai/Qwen3.8-Flash-Next-ROCmFP4-FAST-imatrix-GGUF agentionai/Qwen3.8-Flash-Next-ROCmFP4-FAST-imatrix-GGUF · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 / other | 1549 / 75023 / 87780 | 2026-08-28 |
| deepseek-ai/deepseek-coder-1.3b-base deepseek-ai/deepseek-coder-1.3b-base · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 162 / 30126 / 34619 | 2023-10-28 |
| drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-Anchored-Tensors drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-Anchored-Tensors · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | mit / other | 237 / 24131 / 25814 | 2026-08-01 |
| amd/Instella-MoE-16B-A3B-Think amd/Instella-MoE-16B-A3B-Think · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 2920 / 3138 / 373 | 2026-07-23 |
| ApolloRaines/Jenzin-Wuang-Nemotron-30B-A3B-BF16 ApolloRaines/Jenzin-Wuang-Nemotron-30B-A3B-BF16 · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | other | 10491 / 1117 / 1927 / 2406 | 2026-09-02 |
| d4rkninja/tanpo-marketing d4rkninja/tanpo-marketing · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 1963 / 2004 / 481 | 2026-09-16 |
| atakankutalp/LFM2.5-2.6B-heretic atakankutalp/LFM2.5-2.6B-heretic · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 1312 / 1523 / 313 | 2026-09-25 |
| AItonomy/PhAI-IDE-72B AItonomy/PhAI-IDE-72B · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | other | 1053 / 1098 / 1659 / 395 | 2026-09-15 |
| 0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored 0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 / other | 1021 / 12829 / 53197 / 594 | 2026-08-31 |
| aiuser3993/GLM-Edge-1.5B-Chat-Greek aiuser3993/GLM-Edge-1.5B-Chat-Greek · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 1015 / 211 / 306 | 2026-08-28 |
| Blackfrost-AI/CYBER-FROST-3.8-BF16 Blackfrost-AI/CYBER-FROST-3.8-BF16 · Hugging Face models, GGUF quantisations | model 1 quantisation 3 | other | 2990 / 669 / 682 / 861 | 2026-09-15 |
| agastyasridharan/Qwen2.5-3B-Instruct-Sheldon-SFT-v2 agastyasridharan/Qwen2.5-3B-Instruct-Sheldon-SFT-v2 · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 701 / 797 / 829 | 2026-09-09 |
| Akahsizrr/Cyber-Prime-1.1-2.6B Akahsizrr/Cyber-Prime-1.1-2.6B · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 602 / 637 / 768 | 2026-09-23 |
| ecloudtech/Erk-14B ecloudtech/Erk-14B · Hugging Face models, GGUF quantisations | model 2 quantisation 3 | other | 1853 / 206 / 441 / 712 / 729 | 2026-07-19 |
| d4rkninja/tanpo-ops d4rkninja/tanpo-ops · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 374 / 521 / 547 | 2026-09-18 |
| QuixiAI/WizardLM-7B-Uncensored QuixiAI/WizardLM-7B-Uncensored · Hugging Face models | model 2 | other | 527 | 2023-05-04 |
| Dingdust/ContextPilot-8B-heretic Dingdust/ContextPilot-8B-heretic · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | other | 2476 / 470 / 507 / 697 | 2026-09-06 |
| 0xzknw/LFM2.5-2.6B-Heretic-NX-PRIME 0xzknw/LFM2.5-2.6B-Heretic-NX-PRIME · Hugging Face models | model 2 | other | 377 | 2026-08-25 |
| Dingdust/ContextPilot-E4B-heretic Dingdust/ContextPilot-E4B-heretic · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | other | 1443 / 254 / 270 / 4995 | 2026-09-03 |
| byroneverson/glm-4-9b-chat-abliterated byroneverson/glm-4-9b-chat-abliterated · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 262 / 265 / 963 | 2024-08-24 |
| emberian/h-05b-replay emberian/h-05b-replay · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other | 232 / 237 / 397 | 2026-09-06 |
| alibaba-pai/SearchQwen2.5-3B alibaba-pai/SearchQwen2.5-3B · Hugging Face models | model 2 | other | 230 | 2026-08-24 |
| BlackwoodAI/Darkforest BlackwoodAI/Darkforest · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | other | 128 / 1530 / 179 / 823 | 2026-04-22 |
Facets
Licence 8092 of 10489 claims
other69
Task 1343 of 10489 claims
Sources
Hugging Face models primary
The model itself: its licence, its task and how many people fetch it.
model · failing1568
GGUF quantisations high
Which quantisations exist, which is what decides whether it runs on the machine somebody has.
quantisation · current8921