Models you can run yourself
A model, every quantisation somebody made of it, and what each one costs to run.
Only one source knows: GGUF quantisations 2312 · Hugging Face models 67
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
| Property | Hugging Face models | GGUF quantisations |
|---|---|---|
| Licence | cardData.license | cardData.license |
Everything else the sources say is shown side by side, and not compared.
Things
26–50 of 2,435| Thing | Kind | Licence | Downloads | Date |
|---|---|---|---|---|
| OS-Software/Ternary-Bonsai-2-27B-Uncensored-Heretic-GGUF prism-ml/Ternary-Bonsai-2-27B-gguf · GGUF quantisations | quantisation 37 | apache-2.0 | 0 | 2026-09-17 |
| empero-ai/Qwen3.8-4B-Distill-GGUF empero-ai/Qwen3.8-4B-Distill-GGUF · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 / other | 240 / 666186 / 745965 | 2026-08-15 |
| llmfan46/Qwen3.8-27B-Ultra-Uncensored-Heretic-Native-MTP-Preserved-GGUF llmfan46/Qwen3.8-27B-Ultra-Uncensored-Heretic-Native-MTP-Preserved · GGUF quantisations | quantisation 8 | apache-2.0 | 140 | 2026-08-30 |
| black-forest-labs/FLUX.1-schnell black-forest-labs/FLUX.1-schnell · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | apache-2.0 | 0 / 625507 | 2024-07-31 |
| DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-ULTRA-HERETIC-Uncensored · Hugging Face models, GGUF quantisations | model 2 quantisation 7 | apache-2.0 | 1081 / 1127 / 1382 / 1552 / 2002 / 4241 / 484 / 57515 / 658 | 2026-09-09 |
| empero-ai/Qwen3.8-27B-Ridge-GGUF empero-ai/Qwen3.8-27B-Ridge-GGUF · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 2895 / 423614 / 494692 | 2026-08-15 |
| 0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored 0bserverx/RVN-Qwen3.8-Flash-Next-Abliterated-Uncensored · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 / other | 1021 / 12829 / 55865 / 586 | 2026-08-31 |
| allenai/Olmo-3-7B-Instruct allenai/Olmo-3-7B-Instruct · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 | 444964 / 481737 / 741 / 97 | 2025-11-19 |
| EleutherAI/gpt-neox-20b EleutherAI/gpt-neox-20b · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | apache-2.0 | 40 / 476649 | 2022-04-07 |
| nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2 · GGUF quantisations | quantisation 1 | apache-2.0 | 55029 | 2026-09-03 |
| WalkingCat/Soprano-1.1-80M-GGUF · GGUF quantisations | quantisation 1 | apache-2.0 | 54076 | 2026-08-27 |
| abenzerps/Nex-N2.5-mini-GGUF nex-agi/Nex-N2.5-mini · GGUF quantisations | quantisation 16 | apache-2.0 | 0 | 2026-09-08 |
| black-forest-labs/FLUX.2-klein-4B black-forest-labs/FLUX.2-klein-4B · Hugging Face models, GGUF quantisations | model 2 quantisation 3 | apache-2.0 | 0 / 394622 / 405526 / 69 / 70 | 2026-01-14 |
| SC117/Qwen3.8-Flash-Next-GSQ-RCO-abliterated-GGUF ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF · GGUF quantisations | quantisation 6 | apache-2.0 | 0 | 2026-09-17 |
| NANI-Nithin/K2-Horizon-MoVA-36B-A4B-GGUF IFM/K2-Horizon-MoVA-36B-A4B · GGUF quantisations | quantisation 12 | apache-2.0 | 1091 | 2026-09-03 |
| facebook/hubert-large-ls960-ft facebook/hubert-large-ls960-ft · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 17 / 314662 / 319887 | 2022-03-02 |
| dgpl/dgpl-experimental-1 · GGUF quantisations | quantisation 1 | apache-2.0 | 43958 | 2026-09-30 |
| autotrust/JEV-27B-VL autotrust/JEV-27B-VL · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | apache-2.0 | 267 / 308531 | 2026-09-30 |
| DuoNeural/Cyber-Ornith-1.5-9B-OBLITERATED DuoNeural/Cyber-Ornith-1.5-9B-OBLITERATED · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 | 1213 / 1327 / 1349 / 38391 | 2026-09-25 |
| ATH-MaaS/OvisOCR2 ATH-MaaS/OvisOCR2 · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 197760 / 236122 / 2953 | 2026-07-13 |
| slashreboot/athena-class-model-a google/gemma-4-31B-it · GGUF quantisations | quantisation 18 | apache-2.0 | 0 | 2026-09-04 |
| slashreboot/athena-class-model-a unsloth/gemma-4-31B-it · GGUF quantisations | quantisation 2 | apache-2.0 | 23 | 2026-09-14 |
| BDRC/tibetan-ocr BDRC/tibetan-ocr · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 134966 / 222970 / 65 | 2026-08-17 |
| TokenRhythm/NeoHorse-1-4B-GGUF TokenRhythm/NeoHorse-1-4B · GGUF quantisations | quantisation 8 | apache-2.0 | 11844 | 2026-09-07 |
| CohereLabs/cohere-transcribe-03-2026 CohereLabs/cohere-transcribe-03-2026 · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 | 211795 / 213049 / 41 / 78 | 2026-03-24 |
Facets
Licence 8436 of 10897 claims
apache-2.05631
Task 1389 of 10897 claims
Sources
Hugging Face models primary
The model itself: its licence, its task and how many people fetch it.
model · failing1617
GGUF quantisations high
Which quantisations exist, which is what decides whether it runs on the machine somebody has.
quantisation · current9280