Models you can run yourself
A model, every quantisation somebody made of it, and what each one costs to run.
Only one source knows: GGUF quantisations 2358 · Hugging Face models 64
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
| Property | Hugging Face models | GGUF quantisations |
|---|---|---|
| Licence | cardData.license | cardData.license |
Everything else the sources say is shown side by side, and not compared.
Found
1–25 of 2,513| Thing | Kind | Licence | Downloads | Date |
|---|---|---|---|---|
| BAAI/bge-reranker-v2-m3 BAAI/bge-reranker-v2-m3 · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 / mit | 117 / 17019450 / 17069529 / 472 | 2024-03-15 |
| prism-ml/Ternary-Bonsai-2-27B-gguf Qwen/Qwen3.8-27B · GGUF quantisations | quantisation 284 | apache-2.0 | 0 | 2026-08-26 |
| DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU · Hugging Face models, GGUF quantisations | model 2 quantisation 52 | apache-2.0 / cc-by-4.0 / other | 0 / 1015 / 1016 / 10199 / 1088 / 12341 / 12717 / 1379 / 14563 / 1475 / 1476 / 1511 / 1557 / 1559 / 15592 / 1634 / 1683597 / 1719 / 172 / 1724 / 1754 / 1787 / 1790 / 17942 / 1840 / 1893 / 1894 / 1910 / 1961 / 2027 / 2217 / 2268 / 2311 / 2572 / 2638 / 29374 / 3160 / 31718 / 3255 / 3367 / 3375 / 3470 / 3592 / 400 / 438 / 4756 / 506 / 53 / 5401 / 5748 / 6578 / 703 / 904 / 910 | 2026-08-25 |
| farbodtavakkoli/OTel-2.0-LLM-31B-IT farbodtavakkoli/OTel-2.0-LLM-31B-IT · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 2972402 / 4832934 / 6607 | 2026-07-23 |
| answerdotai/ModernBERT-base answerdotai/ModernBERT-base · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 125 / 4244238 / 4270151 | 2024-12-11 |
| ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Qwen/Qwen3.8-Flash-Next · GGUF quantisations | quantisation 127 | apache-2.0 | 0 | 2026-08-26 |
| XHToken/Spark-X2.5-4B-GGUF XHToken/Spark-X2.5-4B · GGUF quantisations | quantisation 34 | apache-2.0 | 0 | 2026-08-28 |
| empero-ai/Qwen3.8-35B-A3B-Distill empero-ai/Qwen3.8-35B-A3B-Distill · Hugging Face models, GGUF quantisations | model 2 quantisation 11 | apache-2.0 | 1514 / 1756 / 19162 / 21029 / 302 / 303921 / 457 / 4825 / 5362 / 5468 / 6658 / 835 / 9929 | 2026-09-16 |
| DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 | 1603479 / 2164143 / 438 / 4679 | 2026-09-01 |
| Alibaba-NLP/gte-reranker-modernbert-base Alibaba-NLP/gte-reranker-modernbert-base · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | apache-2.0 | 1943461 / 33 | 2025-01-20 |
| docling-project/docling-layout-heron docling-project/docling-layout-heron · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 | 0 / 116 / 1640697 / 1731166 | 2025-04-15 |
| openbmb/MiniCPM5-2B-GGUF · GGUF quantisations | quantisation 1 | apache-2.0 | 234613 | 2026-09-05 |
| 0bserverx/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF 0bserverx/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | apache-2.0 | 0 / 1350398 | 2026-08-14 |
| DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored · Hugging Face models, GGUF quantisations | model 2 quantisation 11 | apache-2.0 | 15792 / 1821 / 1917 / 1970 / 203773 / 2165 / 2376 / 2670 / 4073 / 5187 / 613 / 6902 / 969 | 2026-08-31 |
| amazon/chronos-t5-small amazon/chronos-t5-small · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 1064253 / 1226021 / 26 | 2024-02-21 |
| XHToken/Spark-X2.5-1.7B-GGUF XHToken/Spark-X2.5-1.7B · GGUF quantisations | quantisation 13 | apache-2.0 | 1202 | 2026-08-28 |
| ggml-org/Qwen3.8-27B-GGUF ggml-org/Qwen3.8-27B-GGUF · Hugging Face models, GGUF quantisations | model 1 quantisation 3 | apache-2.0 | 1075427 / 1601 / 323 / 631 | 2026-08-14 |
| pottokao/Qwen-Image-2.1-Text-Encoder-Heretic-GGUF pottokao/Qwen-Image-2.1-Text-Encoder-Heretic · GGUF quantisations | quantisation 6 | apache-2.0 | 1421 | 2026-09-20 |
| Accio-Lab/occamy-1.0 Accio-Lab/occamy-1.0 · Hugging Face models, GGUF quantisations | model 2 quantisation 14 | apache-2.0 | 1010 / 1025 / 1532 / 157 / 158222 / 1759 / 2440 / 2697 / 3431 / 4511 / 4652 / 6035 / 6115 / 8117 / 862 / 9135 | 2026-08-13 |
| Comfy-Org/Qwen3-VL Comfy-Org/Qwen3-VL · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 746 / 750080 / 966109 | 2026-06-03 |
| bartowski/orcarouter_Qwen3.8-27B-Uncensored-GGUF orcarouter/Qwen3.8-27B-Uncensored · GGUF quantisations | quantisation 13 | apache-2.0 | 149535 | 2026-08-27 |
| deepseek-ai/DeepSeek-OCR-2 deepseek-ai/DeepSeek-OCR-2 · Hugging Face models | model 2 | apache-2.0 | 847212 | 2026-01-27 |
| mradermacher/Qwen3.8-Flash-Next-Uncensored-GGUF orcarouter/Qwen3.8-Flash-Next-Uncensored · GGUF quantisations | quantisation 22 | apache-2.0 | 0 | 2026-08-27 |
| SC117/Qwen3.8-Flash-Next-GSQ-RCO-abliterated-GGUF ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF · GGUF quantisations | quantisation 6 | apache-2.0 | 0 | 2026-09-17 |
| byteshape/Qwen3.8-27B-GGUF byteshape/Qwen3.8-27B-GGUF · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 318714 / 832621 / 834 | 2026-08-18 |
Facets
Licence 8718 of 11275 claims
apache-2.05832
Task 1427 of 11275 claims
Sources
Hugging Face models primary
The model itself: its licence, its task and how many people fetch it.
model · failing1660
GGUF quantisations high
Which quantisations exist, which is what decides whether it runs on the machine somebody has.
quantisation · current9615