Open LLM Distribution Leaderboard

This is work in progress, there might be inaccuracies. Last updated: 2026-08-05.

Rank Model Avg Downloads / Month Avg Likes / month Params Created Related repos
1 qwen3-6-35b-a3b 27.8M/mo 1.8K/mo 36.0B 2026-04-153mo ago Qwen/Qwen3.6-35B-A3B-FP8
2 qwen3-6-27b 26.1M/mo 2.2K/mo 27.8B 2026-04-213mo ago Qwen/Qwen3.6-27B-FP8
3 qwen3-vl-2b-instruct 22.4M/mo 47.2/mo 2.13B 2025-10-199mo ago Qwen/Qwen3-VL-2B-Instruct
4 gemma-4-26b-a4b-it 19.6M/mo 686.8/mo 26.5B 2026-03-114mo ago google/gemma-4-26B-A4B-it
5 gemma-4-31b-it 18.2M/mo 1.2K/mo 31.3B 2026-03-114mo ago google/gemma-4-31B-it
6 qwen3-0-6b 12.8M/mo 102.1/mo 0.75B 2025-04-271yr 3mo ago Qwen/Qwen3-0.6B
7 qwen3-5-9b 12.3M/mo 522.8/mo 9.65B 2026-02-275mo ago Qwen/Qwen3.5-9B
8 gemma-4-e4b-it 9.7M/mo 468.6/mo 8.00B 2026-03-025mo ago google/gemma-4-E4B-it
9 qwen2-5-1-5b-instruct 9.5M/mo 41.5/mo 1.54B 2024-09-171yr 10mo ago Qwen/Qwen2.5-1.5B-Instruct
10 gemma-4-12b-it 9.1M/mo 3K/mo 12.0B 2026-05-232mo ago google/gemma-4-12B-it
11 gpt-oss-20b 8.4M/mo 483.6/mo 21.5B 2025-08-0411mo ago openai/gpt-oss-20b
12 qwen2-5-7b-instruct 8.2M/mo 72.9/mo 7.62B 2024-09-161yr 10mo ago Qwen/Qwen2.5-7B-Instruct
13 qwen3-5-4b 8.1M/mo 224.6/mo 4.66B 2026-02-275mo ago Qwen/Qwen3.5-4B
14 llama-3-1-8b-instruct 8.1M/mo 282.8/mo 8.03B 2024-07-182yr ago meta-llama/Llama-3.1-8B-Instruct
15 qwen3-8b 8M/mo 118.4/mo 8.19B 2025-04-271yr 3mo ago Qwen/Qwen3-8B
16 deepseek-v4-flash 6.3M/mo 844.7/mo 291B 2026-04-223mo ago deepseek-ai/DeepSeek-V4-Flash
17 qwen3-5-35b-a3b 6.2M/mo 487.6/mo 36.0B 2026-02-245mo ago Qwen/Qwen3.5-35B-A3B
18 qwen2-5-vl-7b-instruct 5.5M/mo 107.9/mo 8.29B 2025-01-261yr 6mo ago Qwen/Qwen2.5-VL-7B-Instruct
19 qwen2-5-vl-3b-instruct 5.5M/mo 41/mo 3.75B 2025-01-261yr 6mo ago Qwen/Qwen2.5-VL-3B-Instruct
20 qwen3-vl-8b-instruct 5.5M/mo 114.6/mo 8.77B 2025-10-119mo ago Qwen/Qwen3-VL-8B-Instruct
21 llama-3-2-1b-instruct 5.5M/mo 80.5/mo 1.24B 2024-09-181yr 10mo ago meta-llama/Llama-3.2-1B-Instruct
22 qwen3-5-27b 5.3M/mo 332.1/mo 27.8B 2026-02-245mo ago Qwen/Qwen3.5-27B
23 qwen3-4b-instruct-2507 5.3M/mo 82.7/mo 4.02B 2025-08-0511mo ago Qwen/Qwen3-4B-Instruct-2507
24 qwen3-4b 5.2M/mo 57.8/mo 4.02B 2025-04-271yr 3mo ago Qwen/Qwen3-4B
25 qwen2-5-3b-instruct 5.1M/mo 31.8/mo 3.09B 2024-09-171yr 10mo ago Qwen/Qwen2.5-3B-Instruct
26 ornith-1-0-35b 4.6M/mo 933.4/mo 34.7B 2026-06-251mo ago deepreinforce-ai/Ornith-1.0-35B-GGUF
27 diffusiongemma-26b-a4b-it 4.6M/mo 698/mo 25.8B 2026-06-091mo ago google/diffusiongemma-26B-A4B-it
28 qwen3-coder-next 4.4M/mo 432.3/mo 79.7B 2026-02-016mo ago Qwen/Qwen3-Coder-Next-FP8
29 gemma-4-e2b-it 4.3M/mo 262.3/mo 5.12B 2026-03-025mo ago google/gemma-4-E2B-it
30 gpt-oss-120b 4.2M/mo 447.1/mo 117B 2025-08-0411mo ago openai/gpt-oss-120b
31 ornith-1-0-9b 4M/mo 452.1/mo 8.95B 2026-06-251mo ago deepreinforce-ai/Ornith-1.0-9B-GGUF
32 qwen3-32b 4M/mo 56.8/mo 32.8B 2025-04-271yr 3mo ago Qwen/Qwen3-32B
33 qwen3-1-7b 3.9M/mo 33.9/mo 2.03B 2025-04-271yr 3mo ago Qwen/Qwen3-1.7B
34 glm-5 3.9M/mo 469.4/mo 754B 2026-02-115mo ago zai-org/GLM-5-FP8
35 kimi-k2-5 3.8M/mo 421.3/mo 1059B 2026-01-017mo ago moonshotai/Kimi-K2.5
36 glm-5-2 3.5M/mo 3.5K/mo 753B 2026-06-161mo ago zai-org/GLM-5.2
37 qwen3-coder-30b-a3b-instruct 3.4M/mo 195.9/mo 30.5B 2025-07-311yr ago Qwen/Qwen3-Coder-30B-A3B-Instruct
38 qwen3-5-0-8b 3.2M/mo 173.3/mo 0.87B 2026-02-285mo ago Qwen/Qwen3.5-0.8B
39 glm-4-7-flash 3.2M/mo 303.2/mo 31.2B 2026-01-196mo ago zai-org/GLM-4.7-Flash
40 nvidia-nemotron-3-super-120b-a12b 3.1M/mo 255.7/mo 67.2B 2026-03-104mo ago nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
41 qwen3-5-122b-a10b 3.1M/mo 240.3/mo 125B 2026-02-245mo ago Qwen/Qwen3.5-122B-A10B
42 qwen3-vl-4b-instruct 3M/mo 59.8/mo 4.44B 2025-10-119mo ago Qwen/Qwen3-VL-4B-Instruct
43 qwen3-vl-30b-a3b-instruct 2.8M/mo 75.7/mo 31.1B 2025-09-3010mo ago Qwen/Qwen3-VL-30B-A3B-Instruct
44 gemma-3-1b-it 2.8M/mo 66.5/mo 1.00B 2025-03-101yr 4mo ago google/gemma-3-1b-it
45 qwen3-14b 2.8M/mo 45.8/mo 14.8B 2025-04-271yr 3mo ago Qwen/Qwen3-14B
46 deepseek-v4-pro 2.8M/mo 1.6K/mo 1599B 2026-04-223mo ago deepseek-ai/DeepSeek-V4-Pro
47 nemotron-3-nano-omni-30b-a3b-reasoning 2.6M/mo 190.1/mo 18.3B 2026-04-243mo ago nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4
48 deepseek-v3-2 2.6M/mo 180.2/mo 685B 2025-12-018mo ago deepseek-ai/DeepSeek-V3.2
49 qwen2-5-0-5b-instruct 2.6M/mo 30.6/mo 0.49B 2024-09-161yr 10mo ago Qwen/Qwen2.5-0.5B-Instruct
50 qwen2-5-14b-instruct 2.6M/mo 17.4/mo 14.8B 2024-09-161yr 10mo ago Qwen/Qwen2.5-14B-Instruct
51 llama-3-2-1b 2.5M/mo 111.9/mo 1.24B 2024-09-181yr 10mo ago meta-llama/Llama-3.2-1B
52 deepseek-r1 2.5M/mo 733.6/mo 685B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1
53 kimi-k2-6 2.5M/mo 439.5/mo 1059B 2026-04-143mo ago moonshotai/Kimi-K2.6
54 nvidia-nemotron-3-nano-30b-a3b 2.4M/mo 166.6/mo 31.6B 2025-12-048mo ago nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
55 qwen3-5-397b-a17b 2.4M/mo 334.3/mo 403B 2026-02-185mo ago Qwen/Qwen3.5-397B-A17B-FP8
56 llama-3-2-3b-instruct 2.3M/mo 118.2/mo 3.21B 2024-09-181yr 10mo ago meta-llama/Llama-3.2-3B-Instruct
57 qwen3-vl-32b-instruct 2.3M/mo 29.3/mo 33.4B 2025-10-199mo ago Qwen/Qwen3-VL-32B-Instruct
58 meta-llama-3-8b-instruct 2.2M/mo 181.1/mo 8.03B 2024-04-172yr 3mo ago meta-llama/Meta-Llama-3-8B-Instruct
59 qwen3-5-2b 2.2M/mo 93.8/mo 2.27B 2026-02-285mo ago Qwen/Qwen3.5-2B
60 qwen2-5-32b-instruct 2.2M/mo 26.1/mo 32.8B 2024-09-171yr 10mo ago Qwen/Qwen2.5-32B-Instruct
61 qwen3-30b-a3b-instruct-2507 2M/mo 106.3/mo 30.5B 2025-07-281yr ago Qwen/Qwen3-30B-A3B-Instruct-2507
62 qwen2-vl-7b-instruct 2M/mo 57.4/mo 8.29B 2024-08-281yr 11mo ago Qwen/Qwen2-VL-7B-Instruct
63 qwen3-tts-12hz-1-7b-base 2M/mo 72.7/mo 1.93B 2026-01-216mo ago Qwen/Qwen3-TTS-12Hz-1.7B-Base
64 qwen2-vl-2b-instruct 2M/mo 22.3/mo 2.21B 2024-08-281yr 11mo ago Qwen/Qwen2-VL-2B-Instruct
65 qwen3-next-80b-a3b-instruct 1.7M/mo 104.5/mo 81.3B 2025-09-0910mo ago Qwen/Qwen3-Next-80B-A3B-Instruct
66 meta-llama-3-8b 1.7M/mo 243.6/mo 8.03B 2024-04-172yr 3mo ago meta-llama/Meta-Llama-3-8B
67 minimax-m2-7 1.7M/mo 332.4/mo 229B 2026-04-093mo ago MiniMaxAI/MiniMax-M2.7
68 qwen3-30b-a3b 1.7M/mo 69.5/mo 30.5B 2025-04-271yr 3mo ago Qwen/Qwen3-30B-A3B
69 qwen3-6-27b-fable-fusion-711-uncensored-heretic-nm-dau-neo-max-mtp 1.6M/mo 1.5K/mo 26.9B 2026-07-1718d ago DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
70 kimi-k2-7-code 1.6M/mo 861.2/mo 1059B 2026-06-111mo ago moonshotai/Kimi-K2.7-Code
71 qwythos-9b-claude-mythos-5-1m 1.5M/mo 1.7K/mo 8.95B 2026-06-191mo ago empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF
72 deepseek-r1-0528 1.5M/mo 174.1/mo 685B 2025-05-281yr 2mo ago deepseek-ai/DeepSeek-R1-0528
73 deepseek-r1-distill-qwen-32b 1.5M/mo 86.4/mo 32.8B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Qwen-32B
74 qwen2-5-coder-7b-instruct 1.5M/mo 51.7/mo 7.62B 2024-09-171yr 10mo ago Qwen/Qwen2.5-Coder-7B-Instruct
75 phi-3-mini-4k-instruct 1.5M/mo 53/mo 3.82B 2024-04-222yr 3mo ago microsoft/Phi-3-mini-4k-instruct
76 gemma-3-12b-it 1.4M/mo 47/mo 12.2B 2025-03-011yr 5mo ago google/gemma-3-12b-it
77 gemma-3-4b-it 1.4M/mo 82.7/mo 4.30B 2025-02-201yr 5mo ago google/gemma-3-4b-it
78 hy3 1.3M/mo 197/mo 299B 2026-07-0728d ago vcruz305/Hy3-GGUF
79 qwen3-4b-base 1.3M/mo 6.3/mo 4.02B 2025-04-281yr 3mo ago Qwen/Qwen3-4B-Base
80 gemma-3-270m 1.3M/mo 88.7/mo 0.27B 2025-08-0511mo ago google/gemma-3-270m
81 gemma-3-27b-it 1.3M/mo 122.9/mo 27.4B 2025-03-011yr 5mo ago google/gemma-3-27b-it
82 deepseek-r1-0528-qwen3-8b 1.2M/mo 82.3/mo 8.19B 2025-05-291yr 2mo ago deepseek-ai/DeepSeek-R1-0528-Qwen3-8B
83 qwen2-5-0-5b 1.2M/mo 19.2/mo 0.49B 2024-09-151yr 10mo ago Qwen/Qwen2.5-0.5B
84 qwen2-1-5b-instruct 1.2M/mo 6.3/mo 1.54B 2024-06-032yr 2mo ago Qwen/Qwen2-1.5B-Instruct
85 llama-3-1-8b 1.2M/mo 95.7/mo 8.03B 2024-07-142yr ago meta-llama/Llama-3.1-8B
86 minimax-m3 1.1M/mo 740.4/mo 440B 2026-06-022mo ago MiniMaxAI/MiniMax-M3-MXFP8
87 deepseek-r1-distill-qwen-1-5b 1.1M/mo 84.1/mo 1.78B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
88 phi-3-5-mini-instruct 1.1M/mo 48.6/mo 3.82B 2024-08-161yr 11mo ago microsoft/Phi-3.5-mini-instruct
89 qwen2-5-coder-14b-instruct 1.1M/mo 10.3/mo 14.8B 2024-11-061yr 8mo ago Qwen/Qwen2.5-Coder-14B-Instruct
90 qwen2-5-coder-32b-instruct 1M/mo 118.1/mo 32.8B 2024-11-061yr 8mo ago Qwen/Qwen2.5-Coder-32B-Instruct
91 gemma-2-2b 1M/mo 27.8/mo 2.61B 2024-07-162yr ago google/gemma-2-2b
92 llama-2-7b 1M/mo 69/mo 6.74B 2023-07-133yr ago meta-llama/Llama-2-7b-hf
93 deepseek-r1-distill-llama-8b 1M/mo 47.3/mo 8.03B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Llama-8B
94 llama-3-1-70b-instruct 1M/mo 40.2/mo 70.6B 2024-07-162yr ago meta-llama/Llama-3.1-70B-Instruct
95 llama-3-1-nemotron-nano-vl-8b-v1 1M/mo 12.9/mo 8.72B 2025-06-031yr 2mo ago nvidia/Llama-3.1-Nemotron-Nano-VL-8B-V1
96 deepseek-v3 998.1K/mo 214.4/mo 685B 2024-12-251yr 7mo ago deepseek-ai/DeepSeek-V3
97 qwen2-7b-instruct 982.6K/mo 26.4/mo 7.62B 2024-06-042yr 2mo ago Qwen/Qwen2-7B-Instruct
98 qwen3-vl-235b-a22b-instruct 979.5K/mo 43.9/mo 236B 2025-09-2210mo ago Qwen/Qwen3-VL-235B-A22B-Instruct
99 llama-3-3-70b-instruct 942.4K/mo 151.7/mo 70.6B 2024-11-261yr 8mo ago meta-llama/Llama-3.3-70B-Instruct
100 qwen2-5-1-5b 920.2K/mo 9.2/mo 1.54B 2024-09-151yr 10mo ago Qwen/Qwen2.5-1.5B
101 llama-3-1-405b 902.2K/mo 44.8/mo 406B 2024-07-162yr ago meta-llama/Llama-3.1-405B
102 qwen2-5-7b 889.4K/mo 13.3/mo 7.62B 2024-09-151yr 10mo ago Qwen/Qwen2.5-7B
103 qwen2-5-72b-instruct 889.2K/mo 46.6/mo 73.0B 2024-09-171yr 10mo ago Qwen/Qwen2.5-72B-Instruct-AWQ
104 florence-2-large 885.4K/mo 71.6/mo 0.78B 2024-06-152yr 1mo ago microsoft/Florence-2-large
105 llama-2-7b-chat 847.5K/mo 131.1/mo 6.74B 2023-07-133yr ago meta-llama/Llama-2-7b-chat-hf
106 qwen3-omni-30b-a3b-instruct 827.8K/mo 92.5/mo 35.3B 2025-09-2010mo ago Qwen/Qwen3-Omni-30B-A3B-Instruct
107 deepseek-v4-flash-0731 817.9K/mo 2.8K/mo 304B 2026-07-314d ago deepseek-ai/DeepSeek-V4-Flash-0731
108 qwen2-5-vl-32b-instruct 815.1K/mo 34.1/mo 33.5B 2025-03-211yr 4mo ago Qwen/Qwen2.5-VL-32B-Instruct
109 llama-3-2-11b-vision-instruct 808.6K/mo 72.3/mo 10.7B 2024-09-181yr 10mo ago meta-llama/Llama-3.2-11B-Vision-Instruct
110 phi-3-5-vision-instruct 804.6K/mo 31.2/mo 4.15B 2024-08-161yr 11mo ago microsoft/Phi-3.5-vision-instruct
111 deepseek-r1-distill-qwen-7b 800.3K/mo 47.1/mo 7.62B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
112 glm-4-6v-flash 777.3K/mo 79.5/mo unknown 2025-12-087mo ago lmstudio-community/GLM-4.6V-Flash-MLX-4bit
113 qwen3-8b-base 767.7K/mo 7.5/mo 8.19B 2025-04-281yr 3mo ago Qwen/Qwen3-8B-Base
114 qwen3-4b-thinking-2507 763.6K/mo 56.3/mo 4.02B 2025-08-0511mo ago Qwen/Qwen3-4B-Thinking-2507
115 minimax-m2-5 752.8K/mo 263.1/mo 229B 2026-02-125mo ago MiniMaxAI/MiniMax-M2.5
116 kimi-k3 748.2K/mo 6K/mo 2780B 2026-06-131mo ago moonshotai/Kimi-K3
117 florence-2-base 733.5K/mo 15.2/mo 0.23B 2024-06-152yr 1mo ago microsoft/Florence-2-base
118 nvidia-nemotron-3-ultra-550b-a55b 678.6K/mo 284/mo 335B 2026-06-032mo ago nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4
119 phi-2 677.5K/mo 110.3/mo 2.78B 2023-12-132yr 7mo ago microsoft/phi-2
120 phi-4 674.1K/mo 115.6/mo 14.7B 2024-12-111yr 7mo ago microsoft/phi-4
121 qwen-agentworld-35b-a3b 672.7K/mo 164/mo 34.7B 2026-06-241mo ago unsloth/Qwen-AgentWorld-35B-A3B-GGUF
122 nvidia-nemotron-3-nano-4b 638.7K/mo 21.2/mo 3.97B 2026-03-074mo ago nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16
123 gemma-2-2b-it 634.7K/mo 62.5/mo 2.61B 2024-07-162yr ago google/gemma-2-2b-it
124 llama-3-2-3b 629K/mo 39.3/mo 3.21B 2024-09-181yr 10mo ago meta-llama/Llama-3.2-3B
125 laguna-s-2-1 626.9K/mo 562.3/mo 118B 2026-07-021mo ago poolside/Laguna-S-2.1-NVFP4
126 qwopus3-6-35b-a3b-coder-mtp 611.1K/mo 181.8/mo 0.45B 2026-06-291mo ago Jackrong/Qwopus3.6-35B-A3B-Coder-MTP-GGUF
127 llama-4-scout-17b-16e-instruct 600.6K/mo 83.8/mo 109B 2025-04-021yr 4mo ago meta-llama/Llama-4-Scout-17B-16E-Instruct
128 qwen2-5-omni-3b 596.7K/mo 22.7/mo 5.54B 2025-04-301yr 3mo ago Qwen/Qwen2.5-Omni-3B
129 chatglm2-6b 587.7K/mo 55/mo unknown 2023-06-243yr 1mo ago zai-org/chatglm2-6b
130 ornith-1-0-397b 575.6K/mo 134.3/mo 397B 2026-06-251mo ago deepreinforce-ai/Ornith-1.0-397B-FP8
131 mistral-7b-instruct-v0-3 567.5K/mo 5.9/mo 7.25B 2024-05-222yr 2mo ago MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF
132 qwen2-0-5b 554.2K/mo 6.5/mo 0.49B 2024-05-312yr 2mo ago Qwen/Qwen2-0.5B
133 gemma-4-31b 539K/mo 102/mo 32.7B 2026-03-124mo ago google/gemma-4-31B
134 qwen3-235b-a22b 536K/mo 84.9/mo 235B 2025-04-271yr 3mo ago Qwen/Qwen3-235B-A22B
135 gemma-4-e4b 527.1K/mo 74.9/mo 8.00B 2026-03-025mo ago google/gemma-4-E4B
136 qwen3guard-gen-0-6b 520.1K/mo 7.4/mo 0.75B 2025-09-2310mo ago Qwen/Qwen3Guard-Gen-0.6B
137 qwen2-5-vl-72b-instruct 512.8K/mo 39.4/mo 73.4B 2025-01-271yr 6mo ago Qwen/Qwen2.5-VL-72B-Instruct
138 qwythos-9b-v2 488.2K/mo 234/mo 8.95B 2026-07-0926d ago empero-ai/Qwythos-9B-v2-GGUF
139 phi-4-mini-instruct 478.3K/mo 46/mo 3.84B 2025-02-191yr 5mo ago microsoft/Phi-4-mini-instruct
140 deepseek-r1-distill-qwen-14b 471.8K/mo 36.1/mo 14.8B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
141 inkling 471.4K/mo 217/mo 947B 2026-07-1421d ago unsloth/inkling-GGUF
142 qwopus3-6-27b-coder 469.9K/mo 2.5/mo 26.9B 2026-06-291mo ago maci0/Qwopus3.6-27B-Coder-NVFP4
143 deepseek-v4-flash-dspark 467.4K/mo 190.1/mo 165B 2026-06-271mo ago deepseek-ai/DeepSeek-V4-Flash-DSpark
144 deepseek-v3-0324 465.9K/mo 192.4/mo 685B 2025-03-241yr 4mo ago deepseek-ai/DeepSeek-V3-0324
145 qwen3-6-40b-claude-4-6-opus-deckard-heretic-uncensored-thinking-neo-code-di-imatrix-max 456.8K/mo 216.9/mo 39.1B 2026-05-013mo ago DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF
146 meta-llama-3-1-8b-instruct 443.1K/mo 5.4/mo 8.03B 2024-07-192yr ago hugging-quants/Meta-Llama-3.1-8B-Instruct-AWQ-INT4
147 phi-3-mini-128k-instruct 437K/mo 62.1/mo 3.82B 2024-04-222yr 3mo ago microsoft/Phi-3-mini-128k-instruct
148 mage-vl 435.8K/mo 255/mo 4.74B 2026-07-2510d ago microsoft/Mage-VL
149 bonsai-27b-mlx-1bit 434.5K/mo 194.1/mo 1.72B 2026-07-041mo ago prism-ml/Bonsai-27B-mlx-1bit
150 llama-2-13b-chat 431.5K/mo 30.5/mo 13.0B 2023-07-133yr ago meta-llama/Llama-2-13b-chat-hf
151 ternary-bonsai-27b-mlx-2bit 430.8K/mo 158/mo 2.56B 2026-07-041mo ago prism-ml/Ternary-Bonsai-27B-mlx-2bit
152 nemotron-labs-diffusion-8b-base 430.6K/mo 1.1/mo 8.49B 2026-01-146mo ago nvidia/Nemotron-Labs-Diffusion-8B-Base
153 gemma-2-9b-it 430K/mo 34.6/mo 9.24B 2024-06-242yr 1mo ago google/gemma-2-9b-it
154 glm-4-5-air 425.7K/mo 56.6/mo 110B 2025-07-201yr ago zai-org/GLM-4.5-Air
155 gemma-4-26b-a4b 423.9K/mo 75.9/mo 26.5B 2026-03-124mo ago google/gemma-4-26B-A4B
156 bielik-11b-v3-0-instruct 408K/mo 9.7/mo 11.2B 2025-11-078mo ago speakleash/Bielik-11B-v3.0-Instruct
157 deepseek-coder-v2-lite-instruct 404.2K/mo 25/mo 15.7B 2024-06-142yr 1mo ago deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct
158 locateanything-3b 401.5K/mo 558.7/mo 3.83B 2026-03-025mo ago nvidia/LocateAnything-3B
159 qwen3-30b-a3b-thinking-2507 400.3K/mo 32.1/mo 30.5B 2025-07-291yr ago Qwen/Qwen3-30B-A3B-Thinking-2507
160 qwen3-0-6b-base 398.6K/mo 11.8/mo 0.60B 2025-04-281yr 3mo ago Qwen/Qwen3-0.6B-Base
161 molmo2-8b 398.6K/mo 24.8/mo 8.66B 2025-12-147mo ago allenai/Molmo2-8B
162 nvidia-nemotron-nano-9b-v2 397.4K/mo 43.9/mo 8.89B 2025-08-1211mo ago nvidia/NVIDIA-Nemotron-Nano-9B-v2
163 gemmable-4-12b-mtp 393K/mo 30.6/mo 11.9B 2026-06-181mo ago Mia-AiLab/Gemmable-4-12B-MTP-GGUF
164 qwopus3-6-27b-coder-compat-mtp 392.7K/mo 88.1/mo 0.46B 2026-06-201mo ago Jackrong/Qwopus3.6-27B-Coder-Compat-MTP-GGUF
165 qwen2-5-omni-7b 387.9K/mo 118.1/mo 10.7B 2025-03-221yr 4mo ago Qwen/Qwen2.5-Omni-7B
166 phi-tiny-moe-instruct 370.1K/mo 3.1/mo 3.75B 2025-06-231yr 1mo ago microsoft/Phi-tiny-MoE-instruct
167 deepseek-r1-distill-llama-70b 362.9K/mo 43.8/mo 70.6B 2025-01-201yr 6mo ago deepseek-ai/DeepSeek-R1-Distill-Llama-70B
168 olmo-2-0425-1b 361.5K/mo 5.1/mo 1.49B 2025-04-171yr 3mo ago allenai/OLMo-2-0425-1B
169 qwen3-1-7b-base 357.6K/mo 4.9/mo 1.72B 2025-04-281yr 3mo ago Qwen/Qwen3-1.7B-Base
170 t5gemma-s-s-prefixlm 354.9K/mo 0.3/mo 0.31B 2025-06-191yr 1mo ago google/t5gemma-s-s-prefixlm
171 jan-v3-5-4b 354.2K/mo 7/mo 4.41B 2026-03-234mo ago janhq/Jan-v3.5-4B-gguf
172 thinkingcap-qwen3-6-27b 344.3K/mo 197.7/mo 27.3B 2026-07-011mo ago bottlecapai/ThinkingCap-Qwen3.6-27B-GGUF
173 glm-4-1v-9b-thinking 338.2K/mo 59.3/mo 10.3B 2025-06-281yr 1mo ago zai-org/GLM-4.1V-9B-Thinking
174 gemma-2b 333.1K/mo 40.5/mo 2.51B 2024-02-082yr 5mo ago google/gemma-2b
175 qwen3-5-0-8b-base 328.2K/mo 17.4/mo 0.87B 2026-02-285mo ago Qwen/Qwen3.5-0.8B-Base
176 qwen3-5-9b-the-defiant-fable-uncensored-heretic-neo-imatrix-max-mtp 323.1K/mo 265/mo 8.95B 2026-07-1916d ago DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
177 minicpm5-1b-claude-opus-fable5-thinking 317.7K/mo 300.6/mo 1.08B 2026-07-031mo ago GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF
178 devstral-small-2-24b-instruct-2512 314.7K/mo 2.8/mo 4.69B 2025-12-097mo ago mlx-community/Devstral-Small-2-24B-Instruct-2512-4bit
179 qwen2-0-5b-instruct 314K/mo 7.7/mo 0.49B 2024-06-032yr 2mo ago Qwen/Qwen2-0.5B-Instruct
180 qwen3-6-35b-a3b-uncensored-genesis-hermes-v6 308.9K/mo 364/mo 34.7B 2026-07-1322d ago LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF
181 qwen3-vl-235b-a22b-instruct-nvfp4-mlperf-inference-closed-v6-1-fp8-kv 307.1K/mo 1.8/mo 119B 2026-06-151mo ago nvidia/Qwen3-VL-235B-A22B-Instruct-NVFP4-MLPerf-Inference-Closed-V6.1-FP8-KV
182 vntl-llama3-8b-v2 301.1K/mo 0.8/mo 8.03B 2025-01-021yr 7mo ago lmg-anon/vntl-llama3-8b-v2-gguf
183 deepseek-vl2-tiny 301K/mo 12.6/mo 3.37B 2024-12-131yr 7mo ago deepseek-ai/deepseek-vl2-tiny
184 kimi-k2-instruct 298.3K/mo 185.1/mo 1026B 2025-07-111yr ago moonshotai/Kimi-K2-Instruct
185 qwen2-5-3b 298.2K/mo 8.9/mo 3.09B 2024-09-151yr 10mo ago Qwen/Qwen2.5-3B
186 dialogpt-medium 286.8K/mo 8.2/mo unknown 2022-03-024yr 5mo ago microsoft/DialoGPT-medium
187 qwen2-5-coder-1-5b-instruct 283.8K/mo 6/mo 1.54B 2024-09-181yr 10mo ago Qwen/Qwen2.5-Coder-1.5B-Instruct
188 deepseek-v2-lite-chat 283K/mo 5.4/mo 15.7B 2024-05-152yr 2mo ago deepseek-ai/DeepSeek-V2-Lite-Chat
189 qwen2-5-math-1-5b 281.2K/mo 4.9/mo 1.54B 2024-09-161yr 10mo ago Qwen/Qwen2.5-Math-1.5B
190 mimo-v2-5 279.4K/mo 117.9/mo 311B 2026-04-273mo ago XiaomiMiMo/MiMo-V2.5
191 gemma-4-12b 278.5K/mo 285.1/mo 12.0B 2026-05-232mo ago google/gemma-4-12B
192 medgemma-4b-it 277.6K/mo 70.8/mo 4.30B 2025-05-191yr 2mo ago google/medgemma-4b-it
193 qwen2-audio-7b-instruct 273K/mo 22.8/mo 8.40B 2024-07-312yr ago Qwen/Qwen2-Audio-7B-Instruct
194 step3-vl-10b 271.1K/mo 61.4/mo 10.2B 2026-01-136mo ago stepfun-ai/Step3-VL-10B
195 qwen1-5-0-5b-chat 262.8K/mo 3.3/mo 0.62B 2024-01-312yr 6mo ago Qwen/Qwen1.5-0.5B-Chat
196 olmo-3-7b-instruct 262.2K/mo 16.6/mo 7.30B 2025-11-198mo ago allenai/Olmo-3-7B-Instruct
197 minicpm5-1b-claude-opus-fable5-v2-thinking 261.5K/mo 181/mo 1.08B 2026-07-1322d ago GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF
198 gemma-3n-e2b-it 259.1K/mo 22.9/mo 5.44B 2025-06-121yr 1mo ago google/gemma-3n-E2B-it
199 cosmos-reason2-2b 254.7K/mo 18.6/mo 2.44B 2025-12-127mo ago nvidia/Cosmos-Reason2-2B
200 minimax-m2 250.9K/mo 159.4/mo 229B 2025-10-229mo ago MiniMaxAI/MiniMax-M2
201 apertus-8b-instruct-2509 247.9K/mo 41.3/mo 8.05B 2025-08-1311mo ago swiss-ai/Apertus-8B-Instruct-2509
202 medgemma-1-5-4b-it 247K/mo 110.8/mo 4.30B 2026-01-076mo ago google/medgemma-1.5-4b-it
203 qwen2-5-coder-1-5b 244.4K/mo 4.2/mo 1.54B 2024-09-181yr 10mo ago Qwen/Qwen2.5-Coder-1.5B
204 gemma-3-270m-it 242K/mo 49.9/mo 0.27B 2025-07-301yr ago google/gemma-3-270m-it
205 qwen3-vl-8b-thinking 240.9K/mo 22.4/mo 8.77B 2025-10-119mo ago Qwen/Qwen3-VL-8B-Thinking
206 paligemma-3b-mix-224 235.9K/mo 3.8/mo 2.92B 2024-05-122yr 2mo ago google/paligemma-3b-mix-224
207 qwen3-5-2b-base 235.2K/mo 15.7/mo 2.27B 2026-02-285mo ago Qwen/Qwen3.5-2B-Base
208 tinyllama-1-1b-chat-v0-3 230.6K/mo 0.4/mo 1.10B 2023-10-032yr 10mo ago TheBloke/TinyLlama-1.1B-Chat-v0.3-GPTQ
209 mixtral-8x22b-v0-1 227.5K/mo 2.8/mo 141B 2024-04-102yr 3mo ago MaziyarPanahi/Mixtral-8x22B-v0.1-GGUF
210 deepseek-v3-1 227.5K/mo 72.1/mo 685B 2025-08-2111mo ago deepseek-ai/DeepSeek-V3.1
211 meta-llama-3-70b-instruct 227.2K/mo 55.1/mo 70.6B 2024-04-172yr 3mo ago meta-llama/Meta-Llama-3-70B-Instruct
212 step-3-5-flash 225.4K/mo 136.8/mo 199B 2026-02-016mo ago stepfun-ai/Step-3.5-Flash
213 llama-guard-3-8b 223.8K/mo 12.8/mo 8.03B 2024-07-222yr ago meta-llama/Llama-Guard-3-8B
214 qwen2-5-coder-7b 221.8K/mo 7.1/mo 7.62B 2024-09-161yr 10mo ago Qwen/Qwen2.5-Coder-7B
215 kimi-vl-a3b-instruct 218.7K/mo 17.5/mo 16.4B 2025-04-091yr 3mo ago moonshotai/Kimi-VL-A3B-Instruct
216 qwen3-5-9b-base 213.6K/mo 18.5/mo 9.65B 2026-02-265mo ago Qwen/Qwen3.5-9B-Base
217 hyperclovax-seed-text-instruct-1-5b 209.3K/mo 0.3/mo 1.81B 2025-04-241yr 3mo ago rippertnt/HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF
218 gemma-2-27b-it 208K/mo 22.4/mo 27.2B 2024-06-242yr 1mo ago google/gemma-2-27b-it
219 agents-a1 206.5K/mo 1.8/mo 18.9B 2026-07-011mo ago r0b0tlab/Agents-A1-NVFP4
220 qwq-32b 203K/mo 174.1/mo 32.8B 2025-03-051yr 4mo ago Qwen/QwQ-32B
221 meta-llama-3-1-70b-instruct 201.6K/mo 4.5/mo 70.6B 2024-07-192yr ago hugging-quants/Meta-Llama-3.1-70B-Instruct-AWQ-INT4
222 cosmos-reason2-8b 198.5K/mo 26.9/mo 8.77B 2025-12-127mo ago nvidia/Cosmos-Reason2-8B
223 smollm-1-7b-instruct-quantized-w4a16 185K/mo 0/mo 1.84B 2024-08-231yr 11mo ago nm-testing/SmolLM-1.7B-Instruct-quantized.w4a16
224 qwen3-5-4b-base 183.9K/mo 15.2/mo 4.66B 2026-02-275mo ago Qwen/Qwen3.5-4B-Base
225 deepseek-v2-lite 177.4K/mo 6.9/mo 15.7B 2024-05-152yr 2mo ago deepseek-ai/DeepSeek-V2-Lite
226 qwen3-coder-480b-a35b-instruct 172.2K/mo 12.8/mo 480B 2025-07-221yr ago Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8
227 qwen2-5-14b 171.8K/mo 6.9/mo 14.8B 2024-09-151yr 10mo ago Qwen/Qwen2.5-14B
228 hermes-3-llama-3-1-8b 168K/mo 19.7/mo 8.03B 2024-07-282yr ago NousResearch/Hermes-3-Llama-3.1-8B
229 qwen2-5-math-7b 162.7K/mo 5.1/mo 7.62B 2024-09-161yr 10mo ago Qwen/Qwen2.5-Math-7B
230 qwen3-5-27b-claude-4-6-opus-reasoning-distilled-v2 150.1K/mo 3.3/mo 27.8B 2026-03-304mo ago QuantTrio/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2-AWQ
231 hunyuan-mt-7b 148.8K/mo 0.5/mo 7.50B 2025-09-0510mo ago Mungert/Hunyuan-MT-7B-GGUF
232 translategemma-4b-it 148.4K/mo 121.7/mo 4.97B 2026-01-126mo ago google/translategemma-4b-it
233 qwen3-6-35b-a3b-heretic 147.4K/mo 16.7/mo 20.8B 2026-04-173mo ago AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4
234 gemma-4-12b-heretic-abliterated 141.7K/mo 4/mo 11.9B 2026-06-051mo ago culturerevolt/gemma-4-12b-heretic-abliterated-GGUF
235 meta-llama-3-70b 141K/mo 31.8/mo 70.6B 2024-04-172yr 3mo ago meta-llama/Meta-Llama-3-70B
236 qwen3-next-80b-a3b-thinking 141K/mo 2.1/mo 83.8B 2025-09-1210mo ago cyankiwi/Qwen3-Next-80B-A3B-Thinking-AWQ-4bit
237 step-3-7-flash 140.9K/mo 178.7/mo 201B 2026-05-232mo ago stepfun-ai/Step-3.7-Flash
238 qwen3-omni-30b-a3b-thinking 139.3K/mo 29.6/mo 31.7B 2025-09-1510mo ago Qwen/Qwen3-Omni-30B-A3B-Thinking
239 deepseek-v3-2-exp 138.5K/mo 97.6/mo 685B 2025-09-2910mo ago deepseek-ai/DeepSeek-V3.2-Exp
240 phi-3-vision-128k-instruct 135.3K/mo 36.7/mo 4.15B 2024-05-192yr 2mo ago microsoft/Phi-3-vision-128k-instruct
241 lfm2-5-1-2b-instruct 135.2K/mo 29.3/mo 1.17B 2026-01-046mo ago LiquidAI/LFM2.5-1.2B-Instruct-GGUF
242 qwen2-5-coder-3b-instruct 135K/mo 5.8/mo 3.09B 2024-11-061yr 8mo ago Qwen/Qwen2.5-Coder-3B-Instruct
243 mistral-small-24b-instruct-2501 134.1K/mo 1.7/mo 23.6B 2025-01-301yr 6mo ago stelterlab/Mistral-Small-24B-Instruct-2501-AWQ
244 nvidia-nemotron-nano-12b-v2-vl 131.9K/mo 9.3/mo 13.2B 2025-10-219mo ago nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16
245 olmo-3-7b-instruct-sft 129K/mo 0.6/mo 7.30B 2025-11-178mo ago allenai/Olmo-3-7B-Instruct-SFT
246 bonsai-27b 126.2K/mo 11/mo 26.9B 2026-07-305d ago lmstudio-community/Bonsai-27B-GGUF
247 llama-3-3-nemotron-super-49b-v1 125.9K/mo 21.3/mo 49.9B 2025-03-161yr 4mo ago nvidia/Llama-3_3-Nemotron-Super-49B-v1
248 tinyllama-1-1b-chat-v1-0 125.7K/mo 7.9/mo 1.10B 2023-12-312yr 7mo ago TheBloke/TinyLlama-1.1B-Chat-v1.0-GGUF
249 huihui-qwen3-6-35b-a3b-abliterated-fp8-dynamic 124.9K/mo 1.2/mo 36.0B 2026-04-283mo ago coolthor/Huihui-Qwen3.6-35B-A3B-abliterated-FP8-DYNAMIC
250 deepseek-v4-flash-base 122.1K/mo 84.6/mo 292B 2026-04-223mo ago deepseek-ai/DeepSeek-V4-Flash-Base
251 ernie-4-5-vl-28b-a3b-pt 118.3K/mo 7.9/mo 29.4B 2025-06-281yr 1mo ago baidu/ERNIE-4.5-VL-28B-A3B-PT
252 gemma-4-26b-a4b-it-ultra-uncensored-heretic-i1 115.7K/mo 7.2/mo 25.2B 2026-04-133mo ago mradermacher/gemma-4-26B-A4B-it-ultra-uncensored-heretic-i1-GGUF
253 gemma-1-1-2b-it 114.1K/mo 6.1/mo 2.51B 2024-03-262yr 4mo ago google/gemma-1.1-2b-it
254 kwaipilot-kat-coder-v2-5-dev 113.8K/mo 89/mo 34.7B 2026-07-2312d ago bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF
255 qwen1-5-7b 112.1K/mo 1.8/mo 7.72B 2024-01-222yr 6mo ago Qwen/Qwen1.5-7B
256 biogpt 110.8K/mo 6.9/mo unknown 2022-11-203yr 8mo ago microsoft/biogpt
257 deepseek-coder-1-3b-instruct 109.7K/mo 5.3/mo 1.35B 2023-10-292yr 9mo ago deepseek-ai/deepseek-coder-1.3b-instruct
258 mimo-7b-rl 105.5K/mo 18.2/mo 7.83B 2025-04-291yr 3mo ago XiaomiMiMo/MiMo-7B-RL
259 qwen2-1-5b 104.9K/mo 3.9/mo 1.54B 2024-05-312yr 2mo ago Qwen/Qwen2-1.5B
260 qwen3-14b-base 103.7K/mo 3.6/mo 14.8B 2025-04-281yr 3mo ago Qwen/Qwen3-14B-Base
261 qwen2-5-math-1-5b-instruct 103.3K/mo 2.5/mo 1.54B 2024-09-161yr 10mo ago Qwen/Qwen2.5-Math-1.5B-Instruct
262 llada2-0-mini 102.8K/mo 8.4/mo 16.3B 2025-11-258mo ago inclusionAI/LLaDA2.0-mini
263 molmo2-o-7b 98.8K/mo 3.4/mo 7.76B 2025-12-147mo ago allenai/Molmo2-O-7B
264 llama-4-maverick-17b-128e-instruct 97.4K/mo 10.8/mo 402B 2025-04-011yr 4mo ago meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8
265 step3 94.2K/mo 13.6/mo 321B 2025-07-281yr ago stepfun-ai/step3
266 deepseek-coder-6-7b-instruct 93.4K/mo 15.2/mo 6.74B 2023-10-292yr 9mo ago deepseek-ai/deepseek-coder-6.7b-instruct
267 phi-3-5-moe-instruct 92.8K/mo 24.4/mo 41.9B 2024-08-171yr 11mo ago microsoft/Phi-3.5-MoE-instruct
268 mimo-7b-base 92.2K/mo 9/mo 7.83B 2025-04-291yr 3mo ago XiaomiMiMo/MiMo-7B-Base
269 llama-guard-4-12b 92K/mo 7.5/mo 12.0B 2025-04-231yr 3mo ago meta-llama/Llama-Guard-4-12B
270 qwen3-5-35b-a3b-base 91.8K/mo 26.4/mo 36.0B 2026-02-245mo ago Qwen/Qwen3.5-35B-A3B-Base
271 mistral-7b-instruct-v0-2 91.7K/mo 1.6/mo 7.24B 2023-12-112yr 7mo ago TheBloke/Mistral-7B-Instruct-v0.2-AWQ
272 florence-2-base-ft 90.6K/mo 5.6/mo 0.23B 2024-06-152yr 1mo ago microsoft/Florence-2-base-ft
273 nemotron-labs-diffusion-8b 89.8K/mo 11.6/mo 8.49B 2026-03-184mo ago nvidia/Nemotron-Labs-Diffusion-8B
274 olmo-3-1025-7b 89.4K/mo 7.7/mo 7.30B 2025-09-1210mo ago allenai/Olmo-3-1025-7B
275 phi-mini-moe-instruct 88.9K/mo 2.8/mo 7.65B 2025-06-231yr 1mo ago microsoft/Phi-mini-MoE-instruct
276 qwen3-vl-8b-instruct-abliterated 85.7K/mo 0.6/mo 8.19B 2025-11-039mo ago mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF
277 glm-4-5v 81.2K/mo 60.9/mo 108B 2025-08-1011mo ago zai-org/GLM-4.5V
278 sugoi-14b-ultra 80.7K/mo 1.2/mo 14.8B 2025-08-1911mo ago sugoitoolkit/Sugoi-14B-Ultra-GGUF
279 nvidia-nemotron-nano-12b-v2 80.2K/mo 14.3/mo 12.3B 2025-08-2111mo ago nvidia/NVIDIA-Nemotron-Nano-12B-v2
280 paligemma-3b-pt-224 79.8K/mo 20/mo 2.92B 2024-05-122yr 2mo ago google/paligemma-3b-pt-224
281 kimi-vl-a3b-thinking 78.8K/mo 28.4/mo 16.4B 2025-04-091yr 3mo ago moonshotai/Kimi-VL-A3B-Thinking
282 gemma-4-12b-it-qat-q4-0-unquantized 77.1K/mo 11/mo 0.42B 2026-06-042mo ago google/gemma-4-12B-it-qat-q4_0-unquantized-assistant
283 deepseek-coder-7b-instruct-v1-5 76.7K/mo 5.2/mo 6.91B 2024-01-252yr 6mo ago deepseek-ai/deepseek-coder-7b-instruct-v1.5
284 glm-4-5 75.9K/mo 112.5/mo 358B 2025-07-201yr ago zai-org/GLM-4.5
285 qwen3-6-27b-aeon-ultimate-uncensored 73K/mo 25.2/mo 19.1B 2026-04-243mo ago AEON-7/Qwen3.6-27B-AEON-Ultimate-Uncensored-NVFP4
286 minicpm5-1b 72.2K/mo 93.5/mo 1.08B 2026-05-242mo ago openbmb/MiniCPM5-1B-GGUF
287 qwen3-30b-a3b-base 71.6K/mo 4.8/mo 30.5B 2025-04-281yr 3mo ago Qwen/Qwen3-30B-A3B-Base
288 nvlm-d-72b 70.7K/mo 35.1/mo 79.4B 2024-09-301yr 10mo ago nvidia/NVLM-D-72B
289 phi-4-reasoning-plus 67.8K/mo 1/mo 7.84B 2025-09-0510mo ago nvidia/Phi-4-reasoning-plus-NVFP4
290 qwen-72b 67.6K/mo 11.2/mo 72.3B 2023-11-262yr 8mo ago Qwen/Qwen-72B
291 gemma-4-e4b-deckard-heretic 66.9K/mo 0.3/mo 6.20B 2026-04-133mo ago AEON-7/Gemma-4-E4B-DECKARD-HERETIC-NVFP4
292 gpt-oss-safeguard-20b 64.4K/mo 23.2/mo 21.5B 2025-09-1810mo ago openai/gpt-oss-safeguard-20b
293 nemotron-3-5-content-safety 63.3K/mo 16.7/mo 4.30B 2026-05-222mo ago nvidia/Nemotron-3.5-Content-Safety
294 qwopus3-6-35b-a3b-v1-prismascout-blackwell-nvfp4-bf16-vllm-4-75bits 62.8K/mo 3.1/mo 21.2B 2026-05-072mo ago cyburn/Qwopus3.6-35B-A3B-v1-PrismaSCOUT-Blackwell-NVFP4-BF16-vllm-4.75bits
295 llama-3-8b-instruct 62.5K/mo 1.1/mo 8.03B 2024-04-182yr 3mo ago casperhansen/llama-3-8b-instruct-awq
296 qwen3guard-gen-4b 59.1K/mo 5.1/mo 4.41B 2025-09-2310mo ago Qwen/Qwen3Guard-Gen-4B
297 qwen3-6-35b-a3b-uncensored-heretic-native-mtp-preserved 58.1K/mo 2.8/mo 36.0B 2026-05-092mo ago llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GPTQ-Int4
298 qwen1-5-moe-a2-7b 57.3K/mo 7.8/mo 14.3B 2024-02-292yr 5mo ago Qwen/Qwen1.5-MoE-A2.7B
299 olmoe-1b-7b-0924 56.5K/mo 6/mo 6.92B 2024-07-202yr ago allenai/OLMoE-1B-7B-0924
300 command-a-vision-07-2025 52.7K/mo 7.3/mo 112B 2025-07-281yr ago CohereLabs/command-a-vision-07-2025
301 qwen3-coder-next-nvfp4-gb10 50.7K/mo 0.6/mo unknown 2026-04-163mo ago gdubicki/Qwen3-Coder-Next-NVFP4-GB10
302 llama-3-70b-instruct 50.4K/mo 2.6/mo 70.6B 2024-04-182yr 3mo ago casperhansen/llama-3-70b-instruct-awq
303 apertus-70b-instruct-2509-quantized-w4a16 49.6K/mo <0.1/mo 11.3B 2025-09-2110mo ago RedHatAI/Apertus-70B-Instruct-2509-quantized.w4a16
304 ternary-bonsai-8b 49.4K/mo 36.9/mo 8.19B 2026-04-183mo ago prism-ml/Ternary-Bonsai-8B-gguf
305 gemma-4-e4b-agentic-opus-reasoning-geminicli 46K/mo 7.8/mo unknown 2026-04-053mo ago deadbydawn101/gemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-mlx-4bit
306 olmoe-1b-7b-0125-instruct 45.9K/mo 3.7/mo 6.92B 2025-01-271yr 6mo ago allenai/OLMoE-1B-7B-0125-Instruct
307 llama-guard-4-12b-quantized-w4a16 43.7K/mo 0/mo 4.14B 2026-02-165mo ago RedHatAI/Llama-Guard-4-12B-quantized.w4a16
308 wildguard 43.6K/mo 2.2/mo 7.25B 2024-06-152yr 1mo ago allenai/wildguard
309 qwen1-5-moe-a2-7b-chat-quantized-w4a16 39.4K/mo <0.1/mo 14.4B 2025-02-241yr 5mo ago nm-testing/Qwen1.5-MoE-A2.7B-Chat-quantized.w4a16
310 llama-3-2-90b-vision-instruct 38.6K/mo 16/mo 88.6B 2024-09-191yr 10mo ago meta-llama/Llama-3.2-90B-Vision-Instruct
311 ui-tars-1-5-7b 38.5K/mo 1.8/mo 7.62B 2025-04-171yr 3mo ago mradermacher/UI-TARS-1.5-7B-GGUF
312 gpt-oss-puzzle-88b 38.1K/mo 21.6/mo 90.8B 2026-03-254mo ago nvidia/gpt-oss-puzzle-88B
313 paligemma2-3b-ft-docci-448 33.2K/mo 0.7/mo 3.03B 2024-11-211yr 8mo ago google/paligemma2-3b-ft-docci-448
314 internvl3-78b 32.4K/mo 0.7/mo unknown 2025-04-171yr 3mo ago OpenGVLab/InternVL3-78B-AWQ
315 t5gemma-2-270m-270m 32.2K/mo 21.4/mo 0.79B 2025-10-259mo ago google/t5gemma-2-270m-270m
316 exaone-3-5-7-8b-instruct 31.7K/mo 0.9/mo 7.82B 2024-12-011yr 8mo ago LGAI-EXAONE/EXAONE-3.5-7.8B-Instruct-AWQ
317 nvidia-nemotron-nano-12b-v2-vl-nvfp4-qad 31.5K/mo 3/mo 7.70B 2025-10-229mo ago nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-NVFP4-QAD
318 phi-4-multimodal-instruct 30.1K/mo 1.2/mo 4.17B 2025-09-0510mo ago nvidia/Phi-4-multimodal-instruct-NVFP4
319 paligemma-3b-ft-cococap-448 27.1K/mo 0.1/mo 2.92B 2024-05-132yr 2mo ago google/paligemma-3b-ft-cococap-448
320 mistral-large-instruct-2411 19.8K/mo 0.4/mo 123B 2024-11-191yr 8mo ago TechxGenus/Mistral-Large-Instruct-2411-AWQ
321 qwen2-5-coder-14b 19.7K/mo 4.1/mo 14.8B 2024-11-081yr 8mo ago Qwen/Qwen2.5-Coder-14B
322 rex-omni 17.2K/mo 0.4/mo 3.75B 2025-10-319mo ago IDEA-Research/Rex-Omni-AWQ
323 gemma-3-12b-it-abliterated-gptqmodel-4b-128g 17.1K/mo <0.1/mo 12.2B 2025-04-141yr 3mo ago jeffcookio/gemma-3-12b-it-abliterated-gptqmodel-4b-128g
324 ministral-8b-instruct-2410 16.4K/mo <0.1/mo 8.02B 2024-10-241yr 9mo ago PyrTools/Ministral-8B-Instruct-2410-AWQ
325 rocinante-12b-v1-1 12.7K/mo 0/mo 2.79B 2025-07-311yr ago AlfonsoM/Rocinante-12B-v1.1-GPTQ
326 aya-expanse-8b 9.6K/mo <0.1/mo 9.08B 2024-10-261yr 9mo ago Orion-zhen/aya-expanse-8b-AWQ
327 qwen3-4b-instruct-2507-nvfp4a16 8.5K/mo <0.1/mo 2.43B 2025-08-0711mo ago apolloparty/Qwen3-4B-Instruct-2507-NVFP4A16

How This Is Calculated

What counts Public Hugging Face model repos above the minimum download threshold are scanned, then filtered for LLM-ish text, coding, multimodal, and special architecture signals.
Models Related repos are grouped by normalized model name and explicit base-model metadata when available. The main row uses the most-downloaded repo in that group.
Metric scope Grouped metrics sum downloads, all-time downloads, and likes across related lineage repos, including package, quantized, and component repos when metadata or slug evidence links them safely. Representative metrics use only the main repo shown in the Model column.
Default sort Rows are sorted by average downloads per month unless you choose likes per month. Sorting follows the selected metric scope, and visible rank is recalculated after filters.
Rates avg downloads/month = all-time downloads / max(model age, 1 month)
avg likes/month = total likes / max(model age, 1 month)
Coverage Discovery scans configured creators, router models, and popular artifact searches. Valid popular artifacts without a safe grouping target stay as standalone rows. The snapshot omits repos below the configured download floor, currently 100K.
Age filter The default view hides models older than one year. Use Off to include older models from the generated snapshot.
Snapshots The leaderboard updates daily. Each refresh writes a new latest view and keeps timestamped historical snapshots in the bucket. Ages are calculated against the snapshot generation time.
Download counts Hugging Face download counts are public, deduplicated counts over tracked model files, meant to approximate model pulls rather than raw requests for every split model file. The recent downloads field is a recent-window count; downloadsAllTime is the same style of count over all time. This leaderboard uses downloadsAllTime divided by model age for Avg Downloads / Month, while recent downloads appear only in metric tooltips.