Live · run by Zetlyn
Models you can run yourself
A model, every quantisation somebody made of it, and what each one costs to run.
Its entry in the hub · take a copy →
Only one source knows: GGUF quantisations 2526 · Hugging Face models 68
Covers. Every GGUF repository created since 2026-08-23, and the base model behind each one. 2,684 of those bases are named; without a Hugging Face token the run is throttled off after roughly 600 to 750 of them, so the model side of the join is partial and the scope says so on its front page. Excludes. Ollama, whose library names no Hugging Face repository to join on. Benchmarks, which have no machine-readable source.
Compared · what is held against what, and from which column of each source
| Property | Hugging Face models | GGUF quantisations |
|---|---|---|
| Licence | cardData.license | cardData.license |
Everything else the sources say is shown side by side, and not compared.
Things
25 things, from 11,890 claims| Thing | Kind | Licence | Downloads | Date |
|---|---|---|---|---|
| BAAI/bge-small-en-v1.5 BAAI/bge-small-en-v1.5 · Hugging Face models, GGUF quantisations | model 1 quantisation 1 | mit / other | 1020 / 62649609 | 2023-09-12 |
| prism-ml/Ternary-Bonsai-2-27B-gguf Qwen/Qwen3.8-27B · GGUF quantisations | quantisation 277 | apache-2.0 | 0 | 2026-08-26 |
| BAAI/bge-m3 BAAI/bge-m3 · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | mit | 34019592 / 36333889 / 84 / 85 | 2024-01-27 |
| DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU · Hugging Face models, GGUF quantisations | model 2 quantisation 54 | apache-2.0 / cc-by-4.0 / other | 0 / 1015 / 1016 / 10199 / 1088 / 12341 / 12717 / 1379 / 14563 / 1476 / 1511 / 1557 / 1559 / 15592 / 1634 / 1683597 / 1719 / 172 / 1724 / 1754 / 1787 / 1790 / 17942 / 1840 / 1861 / 1878 / 1894 / 1910 / 2027 / 2045 / 2191 / 2217 / 2242 / 2268 / 2311 / 2572 / 2638 / 29374 / 3160 / 31718 / 3255 / 3367 / 3375 / 3470 / 3592 / 406 / 4756 / 506 / 53 / 5401 / 5748 / 6578 / 703 / 904 / 910 | 2026-08-25 |
| Comfy-Org/MiniMax-H3 Comfy-Org/MiniMax-H3 · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | other / unknown | 158 / 21857875 / 23415776 | 2026-07-30 |
| deepseek-ai/DeepSeek-V4.1-Flash deepseek-ai/DeepSeek-V4.1-Flash · Hugging Face models, GGUF quantisations | model 2 quantisation 18 | apache-2.0 / mit | 0 / 1099587 / 12061 / 158 / 15890 / 1693 / 1737 / 19701 / 2157 / 227 / 25057 / 257 / 3550 / 41756 / 5384 / 640577 / 798422 / 9685 | 2026-09-10 |
| abenzerps/Qwen-Image-2.1-Uncensored-GGUF Qwen/Qwen-Image-2.1 · GGUF quantisations | quantisation 83 | apache-2.0 | 0 | 2026-09-20 |
| BAAI/bge-reranker-v2-m3 BAAI/bge-reranker-v2-m3 · Hugging Face models, GGUF quantisations | model 2 quantisation 2 | apache-2.0 / mit | 117 / 17042849 / 17069529 / 472 | 2024-03-15 |
| Serveurperso/YuE2-GGUF m-a-p/YuE2-3B · GGUF quantisations | quantisation 9 | cc-by-nc-4.0 | 128078 | 2026-09-10 |
| Serveurperso/YuE2-GGUF m-a-p/YuE2-Vae · GGUF quantisations | quantisation 1 | cc-by-nc-4.0 | 751004 | 2026-09-12 |
| Serveurperso/YuE2-GGUF m-a-p/SheetSage2 · GGUF quantisations | quantisation 2 | cc-by-nc-4.0 | 17344 | 2026-09-12 |
| Serveurperso/YuE2-GGUF m-a-p/MERT-v2-FullSong · GGUF quantisations | quantisation 1 | cc-by-nc-4.0 | 751004 | 2026-09-12 |
| unsloth/GLM-5.3-GGUF zai-org/GLM-5.3 · GGUF quantisations | quantisation 10 | mit | 0 | 2026-08-28 |
| coqui/XTTS-v2 coqui/XTTS-v2 · Hugging Face models | model 1 | other | 6675119 | 2023-10-31 |
| CMSManhattan/JiRackDeltaNet_27b · GGUF quantisations | quantisation 1 | mit | 552025 | 2026-08-29 |
| farbodtavakkoli/OTel-2.0-LLM-31B-IT farbodtavakkoli/OTel-2.0-LLM-31B-IT · Hugging Face models | model 2 | apache-2.0 | 2706024 | 2026-07-23 |
| ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Qwen/Qwen3.8-Flash-Next · GGUF quantisations | quantisation 7 | apache-2.0 | 127682 | 2026-08-28 |
| deepseek-ai/DeepSeek-V4-Flash-0731 deepseek-ai/DeepSeek-V4-Flash-0731 · Hugging Face models | model 2 | mit | 3824182 | 2026-07-31 |
| OBLITERATUS/Ornith-1.5-9B-OBLITERATED ornith-ai/Ornith-1.5-9B · GGUF quantisations | quantisation 1 | mit | 458453 | 2026-08-27 |
| answerdotai/ModernBERT-base answerdotai/ModernBERT-base · Hugging Face models | model 2 | apache-2.0 | 4188388 | 2024-12-11 |
| XHToken/Spark-X2.5-4B-GGUF XHToken/Spark-X2.5-4B · GGUF quantisations | quantisation 2 | apache-2.0 | 214861 | 2026-08-28 |
| SC117/Qwen3.8-Flash-Next-GSQ-RCO-abliterated-GGUF ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF · GGUF quantisations | quantisation 1 | apache-2.0 | 369229 | 2026-10-02 |
| AngelSlim/Hy4-preview-GGUF · GGUF quantisations | quantisation 1 | — | 356956 | 2026-08-28 |
| empero-ai/Qwen3.8-35B-A3B-Distill empero-ai/Qwen3.8-35B-A3B-Distill · Hugging Face models, GGUF quantisations | model 2 quantisation 1 | apache-2.0 | 303921 / 5362 / 6658 | 2026-09-16 |
| datalab-to/chandra-ocr-2 datalab-to/chandra-ocr-2 · Hugging Face models | model 2 | openrail | 1951652 | 2026-03-16 |
Facets
Task 1484 of 11890 claims
Sources
Hugging Face models primary
The model itself: its licence, its task and how many people fetch it.
model · failing1728
GGUF quantisations high
Which quantisations exist, which is what decides whether it runs on the machine somebody has.
quantisation · current10162