BigHugger
Deployability · 202,261 open models

Which models can you actually run?

Every leaderboard ranks models you call over an API. This one ranks models you can put on your own machine — by the runtime that loads them, the format they ship in, whether they're quantised, and whether the licence lets you sell what you build.

Read from the index 2026-09-18

Two different claims sit behind every bar, and they are kept apart everywhere on this page. Declared means a model card named the runtime. By format means the weights are in a container that runtime reads — which says the file will open, not that the architecture is implemented. A model is counted once if either is true. Nothing declares candle, burn or ort; no one writes a Rust runtime on a model card. That gap is the whole reason this page exists.

Reach, by runtime

🦀 burn113,832
mlx11,998
llama.cpp76,048
vllm110,181
091,046182,092 models

declared on the model cardnot declared, but ships a format it reads

Every bar but one is almost entirely light, which is the finding: for most runtimes the evidence is the file, not the card. MLX is the exception — 11,788 cards name it against 471 models shipping an npz, because mlx-community publishes converted weights under a name that says MLX rather than in the format that proves it. It is the one runtime where the card is the better evidence.

The most-used models candle can load

Either signal
182,092
Declared
0 — nobody writes it on a card
By format
182,092
Commercial use ok
111,698
Quantised build
95,655
ModelDownloadsParamsQuantLicenceTerms
sentence-transformers/all-MiniLM-L6-v2
sentence-similarity
256,481,16123MApache-2.0commercial ok
cross-encoder/ms-marco-MiniLM-L6-v2
text-ranking
88,865,13823MApache-2.0commercial ok
BAAI/bge-small-en-v1.5
feature-extraction
64,638,73933MMITcommercial ok
google-bert/bert-base-uncased
fill-mask
47,693,504110MApache-2.0commercial ok
sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2
sentence-similarity
45,865,568118MApache-2.0commercial ok
google-t5/t5-small
translation
25,000,36161MApache-2.0commercial ok
sentence-transformers/all-mpnet-base-v2
sentence-similarity
23,432,230109MApache-2.0commercial ok
amazon/chronos-2
time-series-forecasting
22,779,369119MApache-2.0commercial ok
FacebookAI/xlm-roberta-base
fill-mask
22,235,431279MMITcommercial ok
Qwen/Qwen3-0.6B
text-generation
22,157,967752MApache-2.0commercial ok
Qwen/Qwen3-VL-8B-Instruct
image-text-to-text
18,432,6868.8BApache-2.0commercial ok
BAAI/bge-reranker-v2-m3
text-classification
18,034,248568MApache-2.0commercial ok
timm/mobilenetv3_small_100.lamb_in1k
image-classification
17,261,8133MApache-2.0commercial ok
openai-community/gpt2
text-generation
15,584,259137MMITcommercial ok
nomic-ai/nomic-embed-text-v1.5
sentence-similarity
15,135,151137MApache-2.0commercial ok
Qwen/Qwen3-8B
text-generation
13,103,6528.2BApache-2.0commercial ok
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF
text-generation
12,866,77931B
base model
1–16-bit
26 builds
Apache-2.0commercial ok
timm/efficientnet_b3.ra2_in1k
image-classification
12,701,41312MApache-2.0commercial ok
intfloat/multilingual-e5-small
sentence-similarity
12,377,202118MMITcommercial ok
BAAI/bge-large-en-v1.5
feature-extraction
12,050,258335MMITcommercial ok
Qwen/Qwen3.6-35B-A3B-FP8
image-text-to-text
10,634,07136B
fp8
Apache-2.0commercial ok
BAAI/bge-base-en-v1.5
feature-extraction
10,534,497109MMITcommercial ok
Qwen/Qwen2.5-7B-Instruct
text-generation
9,880,0067.6BApache-2.0commercial ok
google/gemma-4-26B-A4B-it
image-text-to-text
9,463,92226BApache-2.0commercial ok
sentence-transformers/paraphrase-multilingual-mpnet-base-v2
sentence-similarity
9,445,030278MApache-2.0commercial ok
Qwen/Qwen3.5-9B
image-text-to-text
9,341,5079.7BApache-2.0commercial ok
amazon/chronos-bolt-small
time-series-forecasting
9,145,04148MApache-2.0commercial ok
google/gemma-4-31B-it
image-text-to-text
9,138,73131BApache-2.0commercial ok
autogluon/chronos-2
time-series-forecasting
8,970,857119MApache-2.0commercial ok
unsloth/Qwen3.8-27B-GGUF8,856,15028B
base model
1–16-bit
26 builds
Apache-2.0commercial ok
nvidia/Qwen3.6-35B-A3B-NVFP4
text-generation
8,739,08519B8-bit
modelopt
Apache-2.0commercial ok
Qwen/Qwen3-Embedding-0.6B
feature-extraction
8,526,863596MApache-2.0commercial ok
Qwen/Qwen2.5-0.5B-Instruct
text-generation
8,518,031494MApache-2.0commercial ok
laion/clap-htsat-fused
audio-classification
8,077,041154MApache-2.0commercial ok
Qwen/Qwen3.8-27B-FP8
image-text-to-text
7,767,48128B
fp8
Apache-2.0commercial ok
Comfy-Org/z_image_turbo7,757,1256.2B
base model
Apache-2.0commercial ok
FacebookAI/roberta-base
fill-mask
7,750,559125MMITcommercial ok
Qwen/Qwen3.8-27B
image-text-to-text
7,667,55628BApache-2.0commercial ok
farbodtavakkoli/OTel-2.0-LLM-31B-IT
text-generation
7,588,10931BApache-2.0commercial ok
intfloat/multilingual-e5-base
sentence-similarity
7,460,971278MMITcommercial ok
distilbert/distilbert-base-uncased
fill-mask
7,410,66467MApache-2.0commercial ok
Qwen/Qwen2.5-1.5B-Instruct
text-generation
7,206,6771.5BApache-2.0commercial ok
Qwen/Qwen3.5-4B
image-text-to-text
7,060,6594.7BApache-2.0commercial ok
Qwen/Qwen2.5-VL-7B-Instruct
image-text-to-text
7,050,5758.3BApache-2.0commercial ok
intfloat/multilingual-e5-large
feature-extraction
6,966,155560MMITcommercial ok
Qwen/Qwen3-4B
text-generation
6,897,2064.0BApache-2.0commercial ok
autogluon/chronos-2-small
time-series-forecasting
6,877,23628MApache-2.0commercial ok
google/vit-base-patch16-224
image-classification
6,834,32287MApache-2.0commercial ok
openai/whisper-large-v3-turbo
automatic-speech-recognition
6,787,988809MMITcommercial ok
openai/gpt-oss-20b
text-generation
6,697,24721B8-bit
mxfp4
Apache-2.0commercial ok
ibm-granite/granite-embedding-small-english-r2
feature-extraction
6,398,39848MApache-2.0commercial ok
Qwen/Qwen3.6-27B-FP8
image-text-to-text
6,338,91028B
fp8
Apache-2.0commercial ok
FacebookAI/roberta-large
fill-mask
6,335,648355MMITcommercial ok
autogluon/chronos-bolt-small
time-series-forecasting
6,303,56748MApache-2.0commercial ok
Comfy-Org/Wan_2.2_ComfyUI_Repackaged
image-to-video
5,770,623Apache-2.0commercial ok
ornith-ai/Ornith-1.5-9B-GGUF
text-generation
5,320,5139.0B
name
4–16-bit
5 builds
MITcommercial ok
openai/gpt-oss-120b
text-generation
5,316,371117B8-bit
mxfp4
Apache-2.0commercial ok
BAAI/bge-small-zh-v1.5
feature-extraction
5,115,08424MMITcommercial ok
lmstudio-community/Qwen3.8-27B-MLX-4bit
image-text-to-text
4,967,72227B4-bitApache-2.0commercial ok
openai/whisper-large-v3
automatic-speech-recognition
4,850,9781.5BApache-2.0commercial ok

Ordered by downloads, which on these 202,261 models is a real signal — 96% have a non-zero count. Parameter counts are resolved, not just read: the card first, then the base model's card, then the number in the name, then the GGUF file size, which is a per-parameter figure. That sizes 178,356 of the 182,092 models candle can load; the source is printed under every count that did not come from the card itself. Quantisation is read the same way: the card's quant_bits or method where stated, otherwise the GGUF builds the repository actually ships, which is where a range like 2–16-bit and a build count come from.

The licence column resolved in full is on what you're allowed to ship. Everything here is queryable through the API.