BigHugger
Deployability · 202,261 open models

Which models can you actually run?

Every leaderboard ranks models you call over an API. This one ranks models you can put on your own machine — by the runtime that loads them, the format they ship in, whether they're quantised, and whether the licence lets you sell what you build.

Read from the index 2026-09-18

Two different claims sit behind every bar, and they are kept apart everywhere on this page. Declared means a model card named the runtime. By format means the weights are in a container that runtime reads — which says the file will open, not that the architecture is implemented. A model is counted once if either is true. Nothing declares candle, burn or ort; no one writes a Rust runtime on a model card. That gap is the whole reason this page exists.

Reach, by runtime

🦀 burn113,832
mlx11,998
llama.cpp76,048
vllm110,181
091,046182,092 models

declared on the model cardnot declared, but ships a format it reads

Every bar but one is almost entirely light, which is the finding: for most runtimes the evidence is the file, not the card. MLX is the exception — 11,788 cards name it against 471 models shipping an npz, because mlx-community publishes converted weights under a name that says MLX rather than in the format that proves it. It is the one runtime where the card is the better evidence.

The most-used models llama.cpp can load

Either signal
76,048
Declared
973
By format
76,013
Commercial use ok
51,565
Quantised build
72,603
ModelDownloadsParamsQuantLicenceTerms
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF
text-generation
12,866,77931B
base model
1–16-bit
26 builds
Apache-2.0commercial ok
unsloth/Qwen3.8-27B-GGUF8,856,15028B
base model
1–16-bit
26 builds
Apache-2.0commercial ok
ornith-ai/Ornith-1.5-9B-GGUF
text-generation
5,320,5139.0B
name
4–16-bit
5 builds
MITcommercial ok
ornith-ai/Ornith-1.5-35B-A3B-GGUF
text-generation
4,378,77235B
name
4–16-bit
5 builds
MITcommercial ok
audio-cpp/audio.cpp-gguf
text-to-speech
3,800,3501.5B
gguf size
4–32-bit
8 builds
MITcommercial ok
ornith-ai/Ornith-1.0-9B-GGUF
text-generation
3,514,8319.0B
name
4–16-bit
5 builds
MITcommercial ok
cdiamond/Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF
image-text-to-text
3,307,31228B
base model
Apache-2.0commercial ok
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF
image-text-to-text
2,828,41128B
base model
2–16-bit
26 builds
Apache-2.0commercial ok
JonathanColetti/Qwen3.8-27B-Uncensored-GGUF declared
text-generation
2,688,90928B
base model
2–16-bit
8 builds
Apache-2.0commercial ok
mixedbread-ai/mxbai-embed-large-v1
feature-extraction
2,523,873335MApache-2.0commercial ok
HauhauCS/Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF
image-text-to-text
2,387,04228B
base model
2–16-bit
11 builds
Apache-2.0commercial ok
lmstudio-community/Qwen3.8-27B-GGUF2,328,52228B
base model
4–16-bit
4 builds
Apache-2.0commercial ok
ornith-ai/Ornith-1.0-35B-GGUF
text-generation
2,265,45935B
name
4–16-bit
5 builds
MITcommercial ok
antirez/deepseek-v4-gguf
text-generation
2,031,158291B
base model
4-bit
3 builds
MITcommercial ok
0bserverx/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF
text-generation
1,972,78028B
base model
1–16-bit
25 builds
Apache-2.0commercial ok
unsloth/Qwen3.5-9B-GGUF
image-text-to-text
1,596,0489.7B
base model
2–32-bit
24 builds
Apache-2.0commercial ok
handy-computer/parakeet-unified-en-0.6b-gguf
automatic-speech-recognition
1,562,280600M
name
4–32-bit
6 builds
CC-BY-4.0commercial ok
unsloth/Inkling-Small-GGUF
image-text-to-text
1,435,922266B
base model
1–32-bit
24 builds
Apache-2.0commercial ok
ggml-org/gemma-4-E4B-it-GGUF
any-to-any
1,430,5368.0B
base model
4–16-bit
3 builds
Apache-2.0commercial ok
mudler/KAT-Coder-V2.5-Dev-APEX-GGUF1,364,64535B
base model
Apache-2.0commercial ok
ggml-org/Qwen3.8-27B-GGUF
image-text-to-text
1,354,64028B
base model
4–16-bit
4 builds
Apache-2.0commercial ok
DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
image-text-to-text
1,331,9009.7B
base model
2–32-bit
13 builds
Apache-2.0commercial ok
unsloth/gemma-4-12B-it-qat-GGUF
any-to-any
1,328,58912B
base model
4–32-bit
6 builds
Apache-2.0commercial ok
ornith-ai/Ornith-1.5-397B-GGUF
text-generation
1,327,991397B
name
4–16-bit
5 builds
MITcommercial ok
OBLITERATUS/Qwen3.8-27B-OBLITERATED
text-generation
1,295,69728B2–16-bit
8 builds
Apache-2.0commercial ok
unsloth/Qwen3.6-35B-A3B-GGUF
image-text-to-text
1,268,08336B
base model
1–32-bit
25 builds
Apache-2.0commercial ok
unsloth/Qwen3.6-27B-MTP-GGUF
image-text-to-text
1,230,29328B
base model
2–32-bit
24 builds
Apache-2.0commercial ok
unsloth/Qwen3.6-27B-GGUF
image-text-to-text
1,135,64628B
base model
2–32-bit
24 builds
Apache-2.0commercial ok
mudler/ced-gguf
audio-classification
1,074,34286M
base model
8–32-bit
3 builds
Apache-2.0commercial ok
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive
image-text-to-text
1,068,61836B
base model
2–16-bit
12 builds
Apache-2.0commercial ok
DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
image-text-to-text
1,049,58628B
base model
2–32-bit
12 builds
Apache-2.0commercial ok
bartowski/endless-frontier_BigBang-v1-GGUF
image-text-to-text
986,05236B
base model
2–16-bit
28 builds
Apache-2.0commercial ok
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF
image-text-to-text
956,96428B
base model
2–16-bit
5 builds
Apache-2.0commercial ok
handy-computer/cohere-transcribe-03-2026-gguf
automatic-speech-recognition
955,9002.1B
base model
4–16-bit
6 builds
Apache-2.0commercial ok
LocalAI-io/privacy-filter-nemotron-GGUF
token-classification
953,6581.4B
base model
8–16-bit
2 builds
Apache-2.0commercial ok
unsloth/Qwen3.6-35B-A3B-MTP-GGUF
image-text-to-text
948,19036B
base model
1–32-bit
23 builds
Apache-2.0commercial ok
unsloth/Qwen3.5-4B-GGUF
image-text-to-text
923,1544.7B
base model
2–32-bit
24 builds
Apache-2.0commercial ok
DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
image-text-to-text
907,64028B
base model
2–32-bit
13 builds
Apache-2.0commercial ok
bartowski/XYZAILab_XYZ-Aquila-mini-GGUF
image-text-to-text
883,33036B
base model
2–16-bit
28 builds
Apache-2.0commercial ok
HauhauCS/Qwen3.5-9B-Uncensored-HauhauCS-Aggressive867,3899.7B
base model
4–16-bit
4 builds
Apache-2.0commercial ok
nvidia/parakeet-ctc-1.1b
automatic-speech-recognition
841,2261.1B8-bitCC-BY-4.0commercial ok
FINAL-Bench/POCKET-35B-GGUF declared
text-generation
825,83535B
base model
1–4-bit
4 builds
Apache-2.0commercial ok
google/gemma-4-12B-it-qat-q4_0-gguf
any-to-any
793,35512B
base model
4-bitApache-2.0commercial ok
empero-ai/Qwen3.8-4B-Distill-GGUF
text-generation
771,4344.7B
base model
4–16-bit
5 builds
Apache-2.0commercial ok
empero-ai/Qwen3.8-2B-Distill-GGUF
text-generation
749,2032.3B
base model
4–16-bit
5 builds
Apache-2.0commercial ok
QuantStack/Wan2.2-T2V-A14B-GGUF
text-to-video
724,59214B
gguf size
2–8-bit
13 builds
Apache-2.0commercial ok
empero-ai/Qwen3.8-9B-Distill-GGUF
text-generation
699,4459.7B
base model
4–16-bit
5 builds
Apache-2.0commercial ok
google/gemma-4-E4B-it-qat-q4_0-gguf
any-to-any
693,0708.0B
base model
4-bitApache-2.0commercial ok
yuxinlu1/gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF
text-generation
689,63512B
base model
3–16-bit
6 builds
Apache-2.0commercial ok
unsloth/gemma-4-26B-A4B-it-qat-GGUF
image-text-to-text
683,93927B
base model
4–32-bit
6 builds
Apache-2.0commercial ok
unsloth/gemma-4-31B-it-qat-GGUF
image-text-to-text
679,32033B
base model
4–32-bit
6 builds
Apache-2.0commercial ok
unsloth/inkling-GGUF
image-text-to-text
674,848952B
base model
1–16-bit
7 builds
Apache-2.0commercial ok
mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF669,6328.8B
base model
2–16-bit
12 builds
Apache-2.0commercial ok
unsloth/gemma-4-12b-it-GGUF
image-text-to-text
667,08012B
base model
2–32-bit
23 builds
Apache-2.0commercial ok
unsloth/gemma-4-E4B-it-qat-GGUF
any-to-any
661,3708.0B
base model
2–32-bit
7 builds
Apache-2.0commercial ok
z-lab/Qwen3.8-27B-DFlash2-GGUF declared
text-generation
660,90228B
base model
4–16-bit
3 builds
Apache-2.0commercial ok
prism-ml/Ternary-Bonsai-27B-gguf declared
text-generation
650,95528B
base model
2-bit
6 builds
Apache-2.0commercial ok
peculiar-ragdoll/Tiel-Coder-35B-A3B-GGUF-MTP
image-text-to-text
648,76336B
base model
2–16-bit
10 builds
MITcommercial ok
empero-ai/Qwythos-9B-v2-GGUF
image-text-to-text
648,6769.7B
base model
4–16-bit
5 builds
Apache-2.0commercial ok
unsloth/gemma-4-26B-A4B-it-GGUF
image-text-to-text
646,39027B
base model
2–32-bit
22 builds
Apache-2.0commercial ok

Ordered by downloads, which on these 202,261 models is a real signal — 96% have a non-zero count. Parameter counts are resolved, not just read: the card first, then the base model's card, then the number in the name, then the GGUF file size, which is a per-parameter figure. That sizes 76,023 of the 76,048 models llama.cpp can load; the source is printed under every count that did not come from the card itself. Quantisation is read the same way: the card's quant_bits or method where stated, otherwise the GGUF builds the repository actually ships, which is where a range like 2–16-bit and a build count come from.

The licence column resolved in full is on what you're allowed to ship. Everything here is queryable through the API.