1
qwen3-6-35b-a3b
26.7M/mo
1.8K/mo
36.0B
2026-04-15 3mo ago
Qwen/Qwen3.6-35B-A3B-FP8 + 18 more variants
2
qwen3-6-27b
25.6M/mo
2.3K/mo
27.8B
2026-04-21 3mo ago
Qwen/Qwen3.6-27B-FP8 + 21 more variants
3
qwen3-vl-2b-instruct
22.9M/mo
52.2/mo
2.13B
2025-10-19 9mo ago
Qwen/Qwen3-VL-2B-Instruct + 2 more variants
4
gemma-4-26b-a4b-it
19.8M/mo
718.1/mo
26.5B
2026-03-11 4mo ago
google/gemma-4-26B-A4B-it + 16 more variants
5
gemma-4-31b-it
18.1M/mo
1.2K/mo
32.7B
2026-03-11 4mo ago
google/gemma-4-31B-it + 14 more variants
6
qwen3-0-6b
12.7M/mo
102.1/mo
0.75B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-0.6B + 2 more variants
7
qwen3-5-9b
12.1M/mo
533.1/mo
9.65B
2026-02-27 5mo ago
Qwen/Qwen3.5-9B + 7 more variants
8
gemma-4-e4b-it
9.5M/mo
477.9/mo
8.00B
2026-03-02 4mo ago
google/gemma-4-E4B-it + 10 more variants
9
qwen2-5-1-5b-instruct
9.5M/mo
41.5/mo
1.54B
2024-09-17 1yr 10mo ago
Qwen/Qwen2.5-1.5B-Instruct + 2 more variants
10
gemma-4-12b-it
8.8M/mo
3.2K/mo
12.0B
2026-05-23 2mo ago
google/gemma-4-12B-it + 14 more variants
11
gpt-oss-20b
8.4M/mo
488.3/mo
21.5B
2025-08-04 11mo ago
openai/gpt-oss-20b + 3 more variants
12
qwen2-5-7b-instruct
8.1M/mo
71.4/mo
7.62B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-7B-Instruct + 3 more variants
13
llama-3-1-8b-instruct
8.1M/mo
282.1/mo
8.03B
2024-07-18 2yr ago
meta-llama/Llama-3.1-8B-Instruct + 5 more variants
14
qwen3-5-4b
7.9M/mo
228.4/mo
4.66B
2026-02-27 5mo ago
Qwen/Qwen3.5-4B + 3 more variants
15
qwen3-8b
7.8M/mo
118.3/mo
8.19B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-8B + 7 more variants
16
ornith-1-0-35b
6.9M/mo
1.4K/mo
34.7B
2026-06-25 1mo ago
deepreinforce-ai/Ornith-1.0-35B-GGUF + 5 more variants
17
deepseek-v4-flash
6.4M/mo
818.1/mo
158B
2026-04-22 3mo ago
deepseek-ai/DeepSeek-V4-Flash + 6 more variants
18
qwen3-5-35b-a3b
6.2M/mo
502.6/mo
36.0B
2026-02-24 5mo ago
Qwen/Qwen3.5-35B-A3B + 3 more variants
19
ornith-1-0-9b
5.7M/mo
926/mo
8.95B
2026-06-25 1mo ago
deepreinforce-ai/Ornith-1.0-9B-GGUF + 1 more variant
20
qwen2-5-vl-7b-instruct
5.5M/mo
108.5/mo
8.29B
2025-01-26 1yr 6mo ago
Qwen/Qwen2.5-VL-7B-Instruct + 2 more variants
21
qwen2-5-vl-3b-instruct
5.5M/mo
41.2/mo
3.75B
2025-01-26 1yr 6mo ago
Qwen/Qwen2.5-VL-3B-Instruct + 2 more variants
22
qwen3-vl-8b-instruct
5.4M/mo
115.9/mo
8.77B
2025-10-11 9mo ago
Qwen/Qwen3-VL-8B-Instruct + 3 more variants
23
llama-3-2-1b-instruct
5.4M/mo
80.9/mo
1.24B
2024-09-18 1yr 10mo ago
meta-llama/Llama-3.2-1B-Instruct + 5 more variants
24
qwen3-5-27b
5.4M/mo
341.7/mo
27.8B
2026-02-24 5mo ago
Qwen/Qwen3.5-27B + 4 more variants
25
qwen3-4b-instruct-2507
5.3M/mo
83.7/mo
4.02B
2025-08-05 11mo ago
Qwen/Qwen3-4B-Instruct-2507 + 3 more variants
26
qwen3-4b
5.2M/mo
57.7/mo
4.02B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-4B + 4 more variants
27
qwen2-5-3b-instruct
5.1M/mo
32/mo
3.09B
2024-09-17 1yr 10mo ago
Qwen/Qwen2.5-3B-Instruct + 2 more variants
28
diffusiongemma-26b-a4b-it
4.6M/mo
761.6/mo
25.8B
2026-06-09 1mo ago
google/diffusiongemma-26B-A4B-it + 3 more variants
29
qwen3-coder-next
4.4M/mo
442.4/mo
79.7B
2026-02-01 5mo ago
Qwen/Qwen3-Coder-Next-FP8 + 7 more variants
30
gemma-4-e2b-it
4.2M/mo
265.7/mo
5.12B
2026-03-02 4mo ago
google/gemma-4-E2B-it + 9 more variants
31
gpt-oss-120b
4.2M/mo
451.6/mo
120B
2025-08-04 11mo ago
openai/gpt-oss-120b + 1 more variant
32
qwen3-32b
4M/mo
57.3/mo
32.8B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-32B + 2 more variants
33
glm-5
3.9M/mo
482.3/mo
754B
2026-02-11 5mo ago
zai-org/GLM-5-FP8 + 4 more variants
34
qwen3-1-7b
3.8M/mo
34.2/mo
2.03B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-1.7B + 1 more variant
35
kimi-k2-5
3.8M/mo
431.3/mo
1059B
2026-01-01 6mo ago
moonshotai/Kimi-K2.5 + 2 more variants
36
glm-5-2
3.5M/mo
4.5K/mo
381B
2026-06-22 1mo ago
nvidia/GLM-5.2-NVFP4 + 3 more variants
37
qwen3-coder-30b-a3b-instruct
3.2M/mo
197.1/mo
30.5B
2025-07-31 11mo ago
Qwen/Qwen3-Coder-30B-A3B-Instruct + 8 more variants
38
qwen3-5-0-8b
3.2M/mo
177.4/mo
0.87B
2026-02-28 4mo ago
Qwen/Qwen3.5-0.8B + 2 more variants
39
glm-4-7-flash
3.2M/mo
308.3/mo
31.2B
2026-01-19 6mo ago
zai-org/GLM-4.7-Flash + 5 more variants
40
nvidia-nemotron-3-super-120b-a12b
3.1M/mo
235.4/mo
67.2B
2026-03-10 4mo ago
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 + 3 more variants
41
qwen3-5-122b-a10b
2.9M/mo
226.2/mo
125B
2026-02-24 5mo ago
Qwen/Qwen3.5-122B-A10B + 6 more variants
42
deepseek-v3-2
2.8M/mo
185.2/mo
685B
2025-12-01 7mo ago
deepseek-ai/DeepSeek-V3.2 + 1 more variant
43
deepseek-v4-pro
2.8M/mo
1.6K/mo
1599B
2026-04-22 3mo ago
deepseek-ai/DeepSeek-V4-Pro
44
qwen3-vl-30b-a3b-instruct
2.8M/mo
76.8/mo
31.1B
2025-09-30 9mo ago
Qwen/Qwen3-VL-30B-A3B-Instruct + 3 more variants
45
gemma-3-1b-it
2.7M/mo
66.3/mo
1.00B
2025-03-10 1yr 4mo ago
google/gemma-3-1b-it + 1 more variant
46
qwen3-vl-4b-instruct
2.7M/mo
59.7/mo
4.44B
2025-10-11 9mo ago
Qwen/Qwen3-VL-4B-Instruct + 4 more variants
47
qwen3-14b
2.7M/mo
38.6/mo
14.8B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-14B + 5 more variants
48
nemotron-3-nano-omni-30b-a3b-reasoning
2.6M/mo
196.7/mo
18.3B
2026-04-24 3mo ago
nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-NVFP4 + 2 more variants
49
kimi-k2-6
2.6M/mo
458.2/mo
1059B
2026-04-14 3mo ago
moonshotai/Kimi-K2.6 + 1 more variant
50
llama-3-2-1b
2.6M/mo
112.5/mo
1.24B
2024-09-18 1yr 10mo ago
meta-llama/Llama-3.2-1B
51
qwen2-5-14b-instruct
2.5M/mo
17.5/mo
14.8B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-14B-Instruct + 1 more variant
52
qwen2-5-0-5b-instruct
2.5M/mo
30.6/mo
0.49B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-0.5B-Instruct + 1 more variant
53
nvidia-nemotron-3-nano-30b-a3b
2.5M/mo
171/mo
31.6B
2025-12-04 7mo ago
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 + 3 more variants
54
deepseek-r1
2.4M/mo
739.3/mo
685B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1
55
qwen3-5-397b-a17b
2.4M/mo
343.6/mo
403B
2026-02-18 5mo ago
Qwen/Qwen3.5-397B-A17B-FP8 + 3 more variants
56
llama-3-2-3b-instruct
2.3M/mo
118.8/mo
3.21B
2024-09-18 1yr 10mo ago
meta-llama/Llama-3.2-3B-Instruct + 2 more variants
57
qwen3-vl-32b-instruct
2.3M/mo
29.5/mo
33.4B
2025-10-19 9mo ago
Qwen/Qwen3-VL-32B-Instruct + 2 more variants
58
qwen3-5-2b
2.1M/mo
95.9/mo
2.27B
2026-02-28 4mo ago
Qwen/Qwen3.5-2B + 1 more variant
59
qwen2-5-32b-instruct
2.1M/mo
23/mo
32.8B
2024-09-17 1yr 10mo ago
Qwen/Qwen2.5-32B-Instruct + 5 more variants
60
qwen3-30b-a3b-instruct-2507
2M/mo
107.4/mo
30.5B
2025-07-28 1yr ago
Qwen/Qwen3-30B-A3B-Instruct-2507 + 3 more variants
61
qwen2-vl-7b-instruct
2M/mo
57.8/mo
8.29B
2024-08-28 1yr 11mo ago
Qwen/Qwen2-VL-7B-Instruct + 1 more variant
62
qwen2-vl-2b-instruct
2M/mo
22.5/mo
2.21B
2024-08-28 1yr 11mo ago
Qwen/Qwen2-VL-2B-Instruct
63
qwen3-tts-12hz-1-7b-base
2M/mo
73.1/mo
1.93B
2026-01-21 6mo ago
Qwen/Qwen3-TTS-12Hz-1.7B-Base
64
qwen3-next-80b-a3b-instruct
1.8M/mo
106/mo
81.3B
2025-09-09 10mo ago
Qwen/Qwen3-Next-80B-A3B-Instruct + 1 more variant
65
qwythos-9b-claude-mythos-5-1m
1.7M/mo
1.9K/mo
8.95B
2026-06-19 1mo ago
empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF
66
qwen3-30b-a3b
1.7M/mo
70/mo
30.5B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-30B-A3B + 3 more variants
67
meta-llama-3-8b
1.7M/mo
241.4/mo
8.03B
2024-04-17 2yr 3mo ago
meta-llama/Meta-Llama-3-8B
68
meta-llama-3-8b-instruct
1.7M/mo
174.1/mo
8.03B
2024-04-17 2yr 3mo ago
meta-llama/Meta-Llama-3-8B-Instruct
69
kimi-k2-7-code
1.6M/mo
943/mo
1059B
2026-06-11 1mo ago
moonshotai/Kimi-K2.7-Code + 2 more variants
70
deepseek-r1-0528
1.5M/mo
176.1/mo
685B
2025-05-28 1yr 2mo ago
deepseek-ai/DeepSeek-R1-0528 + 1 more variant
71
deepseek-r1-distill-qwen-32b
1.5M/mo
87.2/mo
32.8B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B + 1 more variant
72
phi-3-mini-4k-instruct
1.5M/mo
53.2/mo
3.82B
2024-04-22 2yr 3mo ago
microsoft/Phi-3-mini-4k-instruct + 1 more variant
73
qwen2-5-coder-7b-instruct
1.5M/mo
51.9/mo
7.62B
2024-09-17 1yr 10mo ago
Qwen/Qwen2.5-Coder-7B-Instruct + 4 more variants
74
minimax-m2-7
1.4M/mo
332.3/mo
229B
2026-04-09 3mo ago
MiniMaxAI/MiniMax-M2.7
75
gemma-3-12b-it
1.4M/mo
47.3/mo
12.2B
2025-03-01 1yr 4mo ago
google/gemma-3-12b-it + 1 more variant
76
gemma-3-4b-it
1.4M/mo
83.4/mo
4.30B
2025-02-20 1yr 5mo ago
google/gemma-3-4b-it
77
qwen3-4b-base
1.3M/mo
6.4/mo
4.02B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-4B-Base
78
gemma-3-270m
1.3M/mo
89.6/mo
0.27B
2025-08-05 11mo ago
google/gemma-3-270m
79
hy3
1.3M/mo
193/mo
299B
2026-07-07 23d ago
vcruz305/Hy3-GGUF + 1 more variant
80
gemma-3-27b-it
1.3M/mo
124.8/mo
27.4B
2025-03-01 1yr 4mo ago
google/gemma-3-27b-it + 4 more variants
81
qwen2-5-0-5b
1.2M/mo
19.3/mo
0.49B
2024-09-15 1yr 10mo ago
Qwen/Qwen2.5-0.5B
82
llama-3-1-8b
1.2M/mo
96.2/mo
8.03B
2024-07-14 2yr ago
meta-llama/Llama-3.1-8B + 1 more variant
83
qwen2-1-5b-instruct
1.2M/mo
6.3/mo
1.54B
2024-06-03 2yr 1mo ago
Qwen/Qwen2-1.5B-Instruct
84
minimax-m3
1.1M/mo
780.2/mo
440B
2026-06-02 1mo ago
MiniMaxAI/MiniMax-M3-MXFP8 + 2 more variants
85
deepseek-r1-0528-qwen3-8b
1.1M/mo
79.4/mo
8.19B
2025-05-29 1yr 2mo ago
deepseek-ai/DeepSeek-R1-0528-Qwen3-8B + 2 more variants
86
deepseek-r1-distill-qwen-1-5b
1.1M/mo
84.5/mo
1.78B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
87
phi-3-5-mini-instruct
1.1M/mo
47.6/mo
3.82B
2024-08-16 1yr 11mo ago
microsoft/Phi-3.5-mini-instruct + 2 more variants
88
gemma-2-2b
1.1M/mo
28/mo
2.61B
2024-07-16 2yr ago
google/gemma-2-2b
89
deepseek-r1-distill-llama-8b
1M/mo
47.7/mo
8.03B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Llama-8B
90
llama-3-1-70b-instruct
1M/mo
40.4/mo
70.6B
2024-07-16 2yr ago
meta-llama/Llama-3.1-70B-Instruct + 1 more variant
91
qwen2-5-coder-14b-instruct
1M/mo
10.4/mo
14.8B
2024-11-06 1yr 8mo ago
Qwen/Qwen2.5-Coder-14B-Instruct + 4 more variants
92
llama-3-1-nemotron-nano-vl-8b-v1
1M/mo
13/mo
8.72B
2025-06-03 1yr 1mo ago
nvidia/Llama-3.1-Nemotron-Nano-VL-8B-V1
93
qwen2-5-coder-32b-instruct
999.8K/mo
112.9/mo
32.8B
2024-11-06 1yr 8mo ago
Qwen/Qwen2.5-Coder-32B-Instruct + 2 more variants
94
deepseek-v3
999.5K/mo
214.7/mo
685B
2024-12-25 1yr 7mo ago
deepseek-ai/DeepSeek-V3
95
qwen2-7b-instruct
985.9K/mo
26.6/mo
7.62B
2024-06-04 2yr 1mo ago
Qwen/Qwen2-7B-Instruct
96
ornith-1-0-397b
966.2K/mo
367.3/mo
397B
2026-06-25 1mo ago
deepreinforce-ai/Ornith-1.0-397B-FP8 + 1 more variant
97
qwen3-vl-235b-a22b-instruct
965.7K/mo
44.5/mo
236B
2025-09-22 10mo ago
Qwen/Qwen3-VL-235B-A22B-Instruct + 1 more variant
98
qwen3-6-27b-fable-fusion-711-uncensored-heretic-nm-dau-neo-max-mtp
955.8K/mo
1K/mo
26.9B
2026-07-17 13d ago
DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF
99
llama-3-3-70b-instruct
931.6K/mo
152.3/mo
70.6B
2024-11-26 1yr 8mo ago
meta-llama/Llama-3.3-70B-Instruct + 4 more variants
100
qwen2-5-1-5b
916K/mo
9.1/mo
1.54B
2024-09-15 1yr 10mo ago
Qwen/Qwen2.5-1.5B + 1 more variant
101
llama-3-1-405b
903.5K/mo
45.1/mo
406B
2024-07-16 2yr ago
meta-llama/Llama-3.1-405B + 1 more variant
102
llama-2-7b
899.9K/mo
64.4/mo
6.74B
2023-07-13 3yr ago
meta-llama/Llama-2-7b-hf
103
qwen2-5-7b
891.7K/mo
13.4/mo
7.62B
2024-09-15 1yr 10mo ago
Qwen/Qwen2.5-7B
104
florence-2-large
885.7K/mo
72.1/mo
0.78B
2024-06-15 2yr 1mo ago
microsoft/Florence-2-large
105
qwen2-5-72b-instruct
885.4K/mo
46.8/mo
73.0B
2024-09-17 1yr 10mo ago
Qwen/Qwen2.5-72B-Instruct-AWQ + 3 more variants
106
llama-2-7b-chat
848.3K/mo
131.6/mo
6.74B
2023-07-13 3yr ago
meta-llama/Llama-2-7b-chat-hf
107
qwen3-omni-30b-a3b-instruct
822K/mo
93.7/mo
35.3B
2025-09-20 10mo ago
Qwen/Qwen3-Omni-30B-A3B-Instruct
108
llama-3-2-11b-vision-instruct
813.8K/mo
72.8/mo
10.7B
2024-09-18 1yr 10mo ago
meta-llama/Llama-3.2-11B-Vision-Instruct
109
deepseek-r1-distill-qwen-7b
804.8K/mo
47.4/mo
7.62B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
110
phi-3-5-vision-instruct
800.9K/mo
31.5/mo
4.15B
2024-08-16 1yr 11mo ago
microsoft/Phi-3.5-vision-instruct
111
glm-4-6v-flash
783.8K/mo
81/mo
unknown
2025-12-08 7mo ago
lmstudio-community/GLM-4.6V-Flash-MLX-4bit + 3 more variants
112
qwen2-5-vl-32b-instruct
774.5K/mo
34.6/mo
33.5B
2025-03-21 1yr 4mo ago
Qwen/Qwen2.5-VL-32B-Instruct + 1 more variant
113
qwen3-8b-base
771.5K/mo
7.4/mo
8.19B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-8B-Base
114
qwen3-4b-thinking-2507
765.3K/mo
57.1/mo
4.02B
2025-08-05 11mo ago
Qwen/Qwen3-4B-Thinking-2507 + 1 more variant
115
minimax-m2-5
746.9K/mo
270.7/mo
229B
2026-02-12 5mo ago
MiniMaxAI/MiniMax-M2.5
116
qwen-agentworld-35b-a3b
739.6K/mo
185.1/mo
34.7B
2026-06-24 1mo ago
unsloth/Qwen-AgentWorld-35B-A3B-GGUF
117
florence-2-base
719.1K/mo
15.3/mo
0.23B
2024-06-15 2yr 1mo ago
microsoft/Florence-2-base
118
qwopus3-6-35b-a3b-coder-mtp
694.2K/mo
204.1/mo
0.45B
2026-06-29 1mo ago
Jackrong/Qwopus3.6-35B-A3B-Coder-MTP-GGUF
119
phi-2
677.7K/mo
110.9/mo
2.78B
2023-12-13 2yr 7mo ago
microsoft/phi-2
120
phi-4
672.9K/mo
116.5/mo
14.7B
2024-12-11 1yr 7mo ago
microsoft/phi-4
121
nvidia-nemotron-3-ultra-550b-a55b
651.9K/mo
302.9/mo
335B
2026-06-03 1mo ago
nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4 + 1 more variant
122
gemma-2-2b-it
634.4K/mo
62.7/mo
2.61B
2024-07-16 2yr ago
google/gemma-2-2b-it + 1 more variant
123
nvidia-nemotron-3-nano-4b
630.6K/mo
21.9/mo
3.97B
2026-03-07 4mo ago
nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16
124
llama-3-2-3b
628.9K/mo
39.5/mo
3.21B
2024-09-18 1yr 10mo ago
meta-llama/Llama-3.2-3B
125
llama-4-scout-17b-16e-instruct
597.7K/mo
84.6/mo
109B
2025-04-02 1yr 3mo ago
meta-llama/Llama-4-Scout-17B-16E-Instruct + 1 more variant
126
chatglm2-6b
588.1K/mo
55.3/mo
unknown
2023-06-24 3yr 1mo ago
zai-org/chatglm2-6b
127
qwen2-5-omni-3b
583.9K/mo
22.9/mo
5.54B
2025-04-30 1yr 3mo ago
Qwen/Qwen2.5-Omni-3B
128
qwen2-0-5b
553.7K/mo
6.5/mo
0.49B
2024-05-31 2yr 1mo ago
Qwen/Qwen2-0.5B
129
gemma-4-e4b
541.7K/mo
76.2/mo
8.00B
2026-03-02 4mo ago
google/gemma-4-E4B
130
qwen3-235b-a22b
532.6K/mo
85.5/mo
235B
2025-04-27 1yr 3mo ago
Qwen/Qwen3-235B-A22B + 3 more variants
131
gemma-4-31b
530.5K/mo
104.3/mo
32.7B
2026-03-12 4mo ago
google/gemma-4-31B
132
qwen3guard-gen-0-6b
525.7K/mo
7.5/mo
0.75B
2025-09-23 10mo ago
Qwen/Qwen3Guard-Gen-0.6B
133
qwen2-5-vl-72b-instruct
515.7K/mo
39.7/mo
73.4B
2025-01-27 1yr 6mo ago
Qwen/Qwen2.5-VL-72B-Instruct + 1 more variant
134
phi-4-mini-instruct
478.6K/mo
46.3/mo
3.84B
2025-02-19 1yr 5mo ago
microsoft/Phi-4-mini-instruct
135
qwopus3-6-27b-coder
476.8K/mo
1.9/mo
16.2B
2026-06-29 1mo ago
maci0/Qwopus3.6-27B-Coder-NVFP4
136
qwen3-6-40b-claude-4-6-opus-deckard-heretic-uncensored-thinking-neo-code-di-imatrix-max
468.3K/mo
222.8/mo
39.1B
2026-05-01 2mo ago
DavidAU/Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking-NEO-CODE-Di-IMatrix-MAX-GGUF
137
deepseek-r1-distill-qwen-14b
465K/mo
36.4/mo
14.8B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
138
qwythos-9b-v2
462K/mo
231/mo
8.95B
2026-07-09 21d ago
empero-ai/Qwythos-9B-v2-GGUF
139
deepseek-v3-0324
460.6K/mo
194.2/mo
685B
2025-03-24 1yr 4mo ago
deepseek-ai/DeepSeek-V3-0324
140
deepseek-v4-flash-dspark
440.2K/mo
204.7/mo
165B
2026-06-27 1mo ago
deepseek-ai/DeepSeek-V4-Flash-DSpark
141
phi-3-mini-128k-instruct
438.1K/mo
62.5/mo
3.82B
2024-04-22 2yr 3mo ago
microsoft/Phi-3-mini-128k-instruct
142
qwopus3-6-27b-coder-compat-mtp
435.6K/mo
99/mo
0.46B
2026-06-20 1mo ago
Jackrong/Qwopus3.6-27B-Coder-Compat-MTP-GGUF
143
llama-2-13b-chat
432.7K/mo
30.6/mo
13.0B
2023-07-13 3yr ago
meta-llama/Llama-2-13b-chat-hf
144
nemotron-labs-diffusion-8b-base
432.7K/mo
1.1/mo
8.49B
2026-01-14 6mo ago
nvidia/Nemotron-Labs-Diffusion-8B-Base
145
gemmable-4-12b-mtp
429K/mo
33.5/mo
11.9B
2026-06-18 1mo ago
Mia-AiLab/Gemmable-4-12B-MTP-GGUF
146
glm-4-5-air
428.4K/mo
57.2/mo
110B
2025-07-20 1yr ago
zai-org/GLM-4.5-Air + 1 more variant
147
gemma-4-26b-a4b
423.8K/mo
77.5/mo
26.5B
2026-03-12 4mo ago
google/gemma-4-26B-A4B
148
gemma-2-9b-it
423.7K/mo
33.8/mo
9.24B
2024-06-24 2yr 1mo ago
google/gemma-2-9b-it + 1 more variant
149
inkling
415.3K/mo
209/mo
947B
2026-07-14 16d ago
unsloth/inkling-GGUF + 1 more variant
150
locateanything-3b
411.6K/mo
570.4/mo
3.83B
2026-03-02 4mo ago
nvidia/LocateAnything-3B
151
bielik-11b-v3-0-instruct
407.8K/mo
9.3/mo
11.2B
2025-11-07 8mo ago
speakleash/Bielik-11B-v3.0-Instruct + 1 more variant
152
qwen3-30b-a3b-thinking-2507
404.4K/mo
32.6/mo
30.5B
2025-07-29 1yr ago
Qwen/Qwen3-30B-A3B-Thinking-2507 + 1 more variant
153
molmo2-8b
402.1K/mo
25.4/mo
8.66B
2025-12-14 7mo ago
allenai/Molmo2-8B
154
deepseek-coder-v2-lite-instruct
401.7K/mo
25.1/mo
15.7B
2024-06-14 2yr 1mo ago
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct + 1 more variant
155
qwen3-0-6b-base
395.2K/mo
11.8/mo
0.60B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-0.6B-Base
156
nvidia-nemotron-nano-9b-v2
394.5K/mo
44.5/mo
8.89B
2025-08-12 11mo ago
nvidia/NVIDIA-Nemotron-Nano-9B-v2 + 1 more variant
157
qwen2-5-omni-7b
387.7K/mo
119.2/mo
10.7B
2025-03-22 1yr 4mo ago
Qwen/Qwen2.5-Omni-7B + 1 more variant
158
thinkingcap-qwen3-6-27b
382.6K/mo
209/mo
27.3B
2026-07-01 29d ago
bottlecapai/ThinkingCap-Qwen3.6-27B-GGUF
159
laguna-s-2-1
380.8K/mo
415/mo
118B
2026-07-02 28d ago
poolside/Laguna-S-2.1-NVFP4 + 1 more variant
160
qwen3-6-27b-heretic-uncensored-finetune-neo-code-di-imatrix-max
376.8K/mo
135.7/mo
26.9B
2026-04-29 3mo ago
DavidAU/Qwen3.6-27B-Heretic-Uncensored-FINETUNE-NEO-CODE-Di-IMatrix-MAX-GGUF
161
deepseek-r1-distill-llama-70b
364.7K/mo
44.1/mo
70.6B
2025-01-20 1yr 6mo ago
deepseek-ai/DeepSeek-R1-Distill-Llama-70B + 1 more variant
162
phi-tiny-moe-instruct
363.4K/mo
3/mo
3.75B
2025-06-23 1yr 1mo ago
microsoft/Phi-tiny-MoE-instruct
163
olmo-2-0425-1b
358.9K/mo
5.2/mo
1.49B
2025-04-17 1yr 3mo ago
allenai/OLMo-2-0425-1B
164
jan-v3-5-4b
356.8K/mo
6.8/mo
4.41B
2026-03-23 4mo ago
janhq/Jan-v3.5-4B-gguf
165
t5gemma-s-s-prefixlm
346.5K/mo
0.3/mo
0.31B
2025-06-19 1yr 1mo ago
google/t5gemma-s-s-prefixlm
166
qwen3-1-7b-base
345K/mo
5/mo
1.72B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-1.7B-Base
167
glm-4-1v-9b-thinking
336.8K/mo
60/mo
10.3B
2025-06-28 1yr 1mo ago
zai-org/GLM-4.1V-9B-Thinking
168
qwen3-5-0-8b-base
327.9K/mo
17.8/mo
0.87B
2026-02-28 4mo ago
Qwen/Qwen3.5-0.8B-Base
169
minicpm5-1b-claude-opus-fable5-thinking
325.5K/mo
312/mo
1.08B
2026-07-03 27d ago
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-Thinking-GGUF
170
qwen2-0-5b-instruct
314.5K/mo
7.8/mo
0.49B
2024-06-03 2yr 1mo ago
Qwen/Qwen2-0.5B-Instruct
171
devstral-small-2-24b-instruct-2512
310.4K/mo
2.9/mo
4.69B
2025-12-09 7mo ago
mlx-community/Devstral-Small-2-24B-Instruct-2512-4bit + 1 more variant
172
meta-llama-3-1-8b-instruct
299.7K/mo
3.8/mo
8.03B
2024-07-19 2yr ago
hugging-quants/Meta-Llama-3.1-8B-Instruct-AWQ-INT4
173
kimi-k2-instruct
299.2K/mo
187.5/mo
1026B
2025-07-11 1yr ago
moonshotai/Kimi-K2-Instruct
174
qwen2-5-3b
298.1K/mo
8.9/mo
3.09B
2024-09-15 1yr 10mo ago
Qwen/Qwen2.5-3B
175
vntl-llama3-8b-v2
297K/mo
0.8/mo
8.03B
2025-01-02 1yr 6mo ago
lmg-anon/vntl-llama3-8b-v2-gguf
176
deepseek-vl2-tiny
290.6K/mo
12.7/mo
3.37B
2024-12-13 1yr 7mo ago
deepseek-ai/deepseek-vl2-tiny
177
gemma-4-12b
290.5K/mo
302.7/mo
12.0B
2026-05-23 2mo ago
google/gemma-4-12B
178
dialogpt-medium
287.3K/mo
8.3/mo
unknown
2022-03-02 4yr 4mo ago
microsoft/DialoGPT-medium
179
qwen3-vl-235b-a22b-instruct-nvfp4-mlperf-inference-closed-v6-1-fp8-kv
284.1K/mo
2/mo
119B
2026-06-15 1mo ago
nvidia/Qwen3-VL-235B-A22B-Instruct-NVFP4-MLPerf-Inference-Closed-V6.1-FP8-KV
180
qwen2-5-math-1-5b
279.1K/mo
4.9/mo
1.54B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-Math-1.5B
181
qwen2-5-coder-1-5b-instruct
277.5K/mo
5.9/mo
1.54B
2024-09-18 1yr 10mo ago
Qwen/Qwen2.5-Coder-1.5B-Instruct
182
medgemma-4b-it
275.5K/mo
71.6/mo
4.30B
2025-05-19 1yr 2mo ago
google/medgemma-4b-it
183
step3-vl-10b
274.5K/mo
62.8/mo
10.2B
2026-01-13 6mo ago
stepfun-ai/Step3-VL-10B
184
deepseek-v2-lite-chat
273.6K/mo
5.4/mo
15.7B
2024-05-15 2yr 2mo ago
deepseek-ai/DeepSeek-V2-Lite-Chat
185
qwen2-audio-7b-instruct
271.2K/mo
23/mo
8.40B
2024-07-31 1yr 11mo ago
Qwen/Qwen2-Audio-7B-Instruct
186
mimo-v2-5
264.4K/mo
121.5/mo
311B
2026-04-27 3mo ago
XiaomiMiMo/MiMo-V2.5
187
qwen1-5-0-5b-chat
262.6K/mo
3.3/mo
0.62B
2024-01-31 2yr 5mo ago
Qwen/Qwen1.5-0.5B-Chat
188
gemma-3n-e2b-it
256.6K/mo
23.2/mo
5.44B
2025-06-12 1yr 1mo ago
google/gemma-3n-E2B-it
189
olmo-3-7b-instruct
255.1K/mo
17/mo
7.30B
2025-11-19 8mo ago
allenai/Olmo-3-7B-Instruct
190
medgemma-1-5-4b-it
250.8K/mo
112.5/mo
4.30B
2026-01-07 6mo ago
google/medgemma-1.5-4b-it
191
minicpm5-1b-claude-opus-fable5-v2-thinking
248.6K/mo
176/mo
1.08B
2026-07-13 17d ago
GnLOLot/MiniCPM5-1B-Claude-Opus-Fable5-V2-Thinking-GGUF
192
qwen3-5-9b-the-defiant-fable-uncensored-heretic-neo-imatrix-max-mtp
248.2K/mo
156/mo
8.95B
2026-07-19 11d ago
DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
193
kimi-k3
247.4K/mo
5.7K/mo
2780B
2026-06-13 1mo ago
moonshotai/Kimi-K3
194
apertus-8b-instruct-2509
246.1K/mo
41.7/mo
8.05B
2025-08-13 11mo ago
swiss-ai/Apertus-8B-Instruct-2509
195
gemma-3-270m-it
244.5K/mo
50.6/mo
0.27B
2025-07-30 1yr ago
google/gemma-3-270m-it
196
qwen2-5-coder-1-5b
244.4K/mo
4.2/mo
1.54B
2024-09-18 1yr 10mo ago
Qwen/Qwen2.5-Coder-1.5B
197
qwen3-vl-8b-thinking
243.6K/mo
22.7/mo
8.77B
2025-10-11 9mo ago
Qwen/Qwen3-VL-8B-Thinking
198
cosmos-reason2-2b
240.6K/mo
18.4/mo
2.44B
2025-12-12 7mo ago
nvidia/Cosmos-Reason2-2B
199
paligemma-3b-mix-224
236K/mo
3.8/mo
2.92B
2024-05-12 2yr 2mo ago
google/paligemma-3b-mix-224
200
meta-llama-3-70b-instruct
227.8K/mo
55.5/mo
70.6B
2024-04-17 2yr 3mo ago
meta-llama/Meta-Llama-3-70B-Instruct
201
deepseek-v3-1
227.1K/mo
73.1/mo
685B
2025-08-21 11mo ago
deepseek-ai/DeepSeek-V3.1
202
tinyllama-1-1b-chat-v0-3
225.8K/mo
0.4/mo
1.10B
2023-10-03 2yr 9mo ago
TheBloke/TinyLlama-1.1B-Chat-v0.3-GPTQ + 1 more variant
203
llama-guard-3-8b
224.5K/mo
12.9/mo
8.03B
2024-07-22 2yr ago
meta-llama/Llama-Guard-3-8B
204
qwen2-5-coder-7b
218.4K/mo
7.1/mo
7.62B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-Coder-7B
205
kimi-vl-a3b-instruct
216.7K/mo
17.7/mo
16.4B
2025-04-09 1yr 3mo ago
moonshotai/Kimi-VL-A3B-Instruct
206
qwen3-5-9b-base
215K/mo
19.3/mo
9.65B
2026-02-26 5mo ago
Qwen/Qwen3.5-9B-Base
207
gemma-2-27b-it
208.7K/mo
22.5/mo
27.2B
2024-06-24 2yr 1mo ago
google/gemma-2-27b-it
208
qwq-32b
203.9K/mo
175.5/mo
32.8B
2025-03-05 1yr 4mo ago
Qwen/QwQ-32B
209
hyperclovax-seed-text-instruct-1-5b
201.7K/mo
0.3/mo
1.81B
2025-04-24 1yr 3mo ago
rippertnt/HyperCLOVAX-SEED-Text-Instruct-1.5B-Q4_K_M-GGUF
210
meta-llama-3-1-70b-instruct
201.7K/mo
4.5/mo
70.6B
2024-07-19 2yr ago
hugging-quants/Meta-Llama-3.1-70B-Instruct-AWQ-INT4
211
cosmos-reason2-8b
186.3K/mo
27.5/mo
8.77B
2025-12-12 7mo ago
nvidia/Cosmos-Reason2-8B
212
qwen3-5-2b-base
184.2K/mo
16.2/mo
2.27B
2026-02-28 4mo ago
Qwen/Qwen3.5-2B-Base
213
qwen3-5-4b-base
182.3K/mo
15.5/mo
4.66B
2026-02-27 5mo ago
Qwen/Qwen3.5-4B-Base
214
smollm-1-7b-instruct-quantized-w4a16
177.7K/mo
0/mo
1.84B
2024-08-23 1yr 11mo ago
nm-testing/SmolLM-1.7B-Instruct-quantized.w4a16
215
deepseek-v2-lite
175.7K/mo
6.9/mo
15.7B
2024-05-15 2yr 2mo ago
deepseek-ai/DeepSeek-V2-Lite
216
qwen3-coder-480b-a35b-instruct
171.3K/mo
13/mo
480B
2025-07-22 1yr ago
Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8
217
agents-a1
167.6K/mo
2/mo
18.9B
2026-07-01 29d ago
r0b0tlab/Agents-A1-NVFP4
218
qwen2-5-math-7b
163K/mo
5.2/mo
7.62B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-Math-7B
219
qwen3-6-35b-a3b-uncensored-genesis-hermes-v6
162.4K/mo
241/mo
34.7B
2026-07-13 17d ago
LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V6-GGUF
220
rnj-1-instruct
162.3K/mo
0.6/mo
8.84B
2025-12-07 7mo ago
Doradus-AI/RnJ-1-Instruct-FP8
221
qwen3-5-27b-claude-4-6-opus-reasoning-distilled-v2
151.2K/mo
3.5/mo
27.8B
2026-03-30 4mo ago
QuantTrio/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-v2-AWQ
222
hunyuan-mt-7b
149.1K/mo
0.6/mo
7.50B
2025-09-05 10mo ago
Mungert/Hunyuan-MT-7B-GGUF
223
qwen3-next-80b-a3b-thinking
141.6K/mo
2.2/mo
83.8B
2025-09-12 10mo ago
cyankiwi/Qwen3-Next-80B-A3B-Thinking-AWQ-4bit
224
meta-llama-3-70b
140.6K/mo
32/mo
70.6B
2024-04-17 2yr 3mo ago
meta-llama/Meta-Llama-3-70B
225
step-3-7-flash
139.8K/mo
189.9/mo
201B
2026-05-23 2mo ago
stepfun-ai/Step-3.7-Flash
226
deepseek-v3-2-exp
138.3K/mo
99.1/mo
685B
2025-09-29 10mo ago
deepseek-ai/DeepSeek-V3.2-Exp
227
qwen3-6-35b-a3b-heretic
135.9K/mo
17.5/mo
20.8B
2026-04-17 3mo ago
AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4
228
qwen2-5-coder-3b-instruct
135.2K/mo
5.7/mo
3.09B
2024-11-06 1yr 8mo ago
Qwen/Qwen2.5-Coder-3B-Instruct
229
phi-3-vision-128k-instruct
134.7K/mo
36.9/mo
4.15B
2024-05-19 2yr 2mo ago
microsoft/Phi-3-vision-128k-instruct
230
lfm2-5-1-2b-instruct
134.5K/mo
29/mo
1.17B
2026-01-04 6mo ago
LiquidAI/LFM2.5-1.2B-Instruct-GGUF
231
qwen3-omni-30b-a3b-thinking
134.1K/mo
30/mo
31.7B
2025-09-15 10mo ago
Qwen/Qwen3-Omni-30B-A3B-Thinking
232
mistral-small-24b-instruct-2501
132.7K/mo
1.7/mo
23.6B
2025-01-30 1yr 5mo ago
stelterlab/Mistral-Small-24B-Instruct-2501-AWQ
233
nvidia-nemotron-nano-12b-v2-vl
131.4K/mo
9.5/mo
13.2B
2025-10-21 9mo ago
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16
234
gemma-4-12b-heretic-abliterated
131.1K/mo
3.8/mo
11.9B
2026-06-05 1mo ago
culturerevolt/gemma-4-12b-heretic-abliterated-GGUF
235
huihui-qwen3-6-35b-a3b-abliterated-fp8-dynamic
131K/mo
1/mo
36.0B
2026-04-28 3mo ago
coolthor/Huihui-Qwen3.6-35B-A3B-abliterated-FP8-DYNAMIC
236
olmo-3-7b-instruct-sft
130.1K/mo
0.6/mo
7.30B
2025-11-17 8mo ago
allenai/Olmo-3-7B-Instruct-SFT
237
tinyllama-1-1b-chat-v1-0
125K/mo
8/mo
1.10B
2023-12-31 2yr 6mo ago
TheBloke/TinyLlama-1.1B-Chat-v1.0-GGUF + 1 more variant
238
deepseek-coder-6-7b-instruct
124.3K/mo
15.9/mo
6.74B
2023-10-29 2yr 9mo ago
deepseek-ai/deepseek-coder-6.7b-instruct + 1 more variant
239
llama-3-3-nemotron-super-49b-v1
123K/mo
21.4/mo
49.9B
2025-03-16 1yr 4mo ago
nvidia/Llama-3_3-Nemotron-Super-49B-v1 + 1 more variant
240
ernie-4-5-vl-28b-a3b-pt
116.8K/mo
8/mo
29.4B
2025-06-28 1yr 1mo ago
baidu/ERNIE-4.5-VL-28B-A3B-PT
241
gemma-1-1-2b-it
113.5K/mo
6.2/mo
2.51B
2024-03-26 2yr 4mo ago
google/gemma-1.1-2b-it
242
qwen1-5-7b
112.1K/mo
1.9/mo
7.72B
2024-01-22 2yr 6mo ago
Qwen/Qwen1.5-7B
243
deepseek-v4-flash-base
111.2K/mo
87/mo
292B
2026-04-22 3mo ago
deepseek-ai/DeepSeek-V4-Flash-Base
244
biogpt
110.9K/mo
6.9/mo
unknown
2022-11-20 3yr 8mo ago
microsoft/biogpt
245
gemma-4-26b-a4b-it-ultra-uncensored-heretic-i1
109.1K/mo
6.7/mo
25.2B
2026-04-13 3mo ago
mradermacher/gemma-4-26B-A4B-it-ultra-uncensored-heretic-i1-GGUF
246
qwen2-1-5b
104.7K/mo
3.9/mo
1.54B
2024-05-31 2yr 1mo ago
Qwen/Qwen2-1.5B
247
qwen3-14b-base
103.2K/mo
3.6/mo
14.8B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-14B-Base
248
qwen2-5-math-1-5b-instruct
102.9K/mo
2.5/mo
1.54B
2024-09-16 1yr 10mo ago
Qwen/Qwen2.5-Math-1.5B-Instruct
249
mimo-7b-rl
102.8K/mo
18.4/mo
7.83B
2025-04-29 1yr 3mo ago
XiaomiMiMo/MiMo-7B-RL
250
llada2-0-mini
101.5K/mo
8.6/mo
16.3B
2025-11-25 8mo ago
inclusionAI/LLaDA2.0-mini
251
llama-4-maverick-17b-128e-instruct
97.3K/mo
10.9/mo
402B
2025-04-01 1yr 3mo ago
meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8
252
molmo2-o-7b
96.6K/mo
3.5/mo
7.76B
2025-12-14 7mo ago
allenai/Molmo2-O-7B
253
step3
94.1K/mo
13.7/mo
321B
2025-07-28 1yr ago
stepfun-ai/step3
254
phi-3-5-moe-instruct
92.3K/mo
24.5/mo
41.9B
2024-08-17 1yr 11mo ago
microsoft/Phi-3.5-MoE-instruct
255
nemotron-labs-diffusion-8b
91K/mo
12/mo
8.49B
2026-03-18 4mo ago
nvidia/Nemotron-Labs-Diffusion-8B
256
qwen3-5-35b-a3b-base
90.7K/mo
27.2/mo
36.0B
2026-02-24 5mo ago
Qwen/Qwen3.5-35B-A3B-Base
257
mimo-7b-base
90K/mo
9.1/mo
7.83B
2025-04-29 1yr 3mo ago
XiaomiMiMo/MiMo-7B-Base
258
llama-guard-4-12b
89.7K/mo
7.6/mo
12.0B
2025-04-23 1yr 3mo ago
meta-llama/Llama-Guard-4-12B
259
mistral-7b-instruct-v0-2
89.2K/mo
1.6/mo
7.24B
2023-12-11 2yr 7mo ago
TheBloke/Mistral-7B-Instruct-v0.2-AWQ
260
olmo-3-1025-7b
89.2K/mo
7.6/mo
7.30B
2025-09-12 10mo ago
allenai/Olmo-3-1025-7B
261
florence-2-base-ft
88.9K/mo
5.6/mo
0.23B
2024-06-15 2yr 1mo ago
microsoft/Florence-2-base-ft
262
phi-mini-moe-instruct
87.4K/mo
2.9/mo
7.65B
2025-06-23 1yr 1mo ago
microsoft/Phi-mini-MoE-instruct
263
nvidia-nemotron-nano-12b-v2
81.3K/mo
14.5/mo
12.3B
2025-08-21 11mo ago
nvidia/NVIDIA-Nemotron-Nano-12B-v2
264
glm-4-5v
80.3K/mo
61.7/mo
108B
2025-08-10 11mo ago
zai-org/GLM-4.5V
265
paligemma-3b-pt-224
78.9K/mo
19.9/mo
2.92B
2024-05-12 2yr 2mo ago
google/paligemma-3b-pt-224
266
sugoi-14b-ultra
78.4K/mo
1.1/mo
14.8B
2025-08-19 11mo ago
sugoitoolkit/Sugoi-14B-Ultra-GGUF
267
hy-mt2-1-8b
77.7K/mo
436.6/mo
2.04B
2026-05-11 2mo ago
tencent/Hy-MT2-1.8B
268
kimi-vl-a3b-thinking
77.5K/mo
28.7/mo
16.4B
2025-04-09 1yr 3mo ago
moonshotai/Kimi-VL-A3B-Thinking
269
qwen3-6-27b-aeon-ultimate-uncensored
76.5K/mo
25.8/mo
19.1B
2026-04-24 3mo ago
AEON-7/Qwen3.6-27B-AEON-Ultimate-Uncensored-NVFP4
270
hy-mt2-30b-a3b
75.5K/mo
182.6/mo
30.1B
2026-05-11 2mo ago
tencent/Hy-MT2-30B-A3B
271
glm-4-5
75.1K/mo
114/mo
358B
2025-07-20 1yr ago
zai-org/GLM-4.5
272
gemma-4-12b-it-qat-q4-0-unquantized
74.8K/mo
11.9/mo
0.42B
2026-06-04 1mo ago
google/gemma-4-12B-it-qat-q4_0-unquantized-assistant
273
qwen3-vl-8b-instruct-abliterated
74.4K/mo
0.3/mo
8.19B
2025-11-03 8mo ago
mradermacher/Qwen3-VL-8B-Instruct-abliterated-GGUF
274
deepseek-coder-7b-instruct-v1-5
71.8K/mo
5.2/mo
6.91B
2024-01-25 2yr 6mo ago
deepseek-ai/deepseek-coder-7b-instruct-v1.5
275
qwen3-30b-a3b-base
71.2K/mo
4.8/mo
30.5B
2025-04-28 1yr 3mo ago
Qwen/Qwen3-30B-A3B-Base
276
gemma-4-e4b-deckard-heretic
69.9K/mo
0.3/mo
6.20B
2026-04-13 3mo ago
AEON-7/Gemma-4-E4B-DECKARD-HERETIC-NVFP4
277
nvlm-d-72b
69.8K/mo
35.3/mo
79.4B
2024-09-30 1yr 9mo ago
nvidia/NVLM-D-72B
278
gemma-4-12b-it-abliterated
69.8K/mo
1.7/mo
7.71B
2026-06-07 1mo ago
pekkAi/Gemma-4-12B-it-abliterated-NVFP4
279
granite-4-1-3b
69.7K/mo
2.9/mo
3.40B
2026-04-16 3mo ago
ibm-granite/granite-4.1-3b-GGUF
280
minicpm5-1b
69.1K/mo
98.2/mo
1.08B
2026-05-24 2mo ago
openbmb/MiniCPM5-1B-GGUF
281
phi-4-reasoning-plus
68.8K/mo
1/mo
7.84B
2025-09-05 10mo ago
nvidia/Phi-4-reasoning-plus-NVFP4
282
qwopus3-6-35b-a3b-v1-prismascout-blackwell-nvfp4-bf16-vllm-4-75bits
66.4K/mo
3.2/mo
21.2B
2026-05-07 2mo ago
cyburn/Qwopus3.6-35B-A3B-v1-PrismaSCOUT-Blackwell-NVFP4-BF16-vllm-4.75bits
283
gpt-oss-safeguard-20b
64.8K/mo
23.5/mo
21.5B
2025-09-18 10mo ago
openai/gpt-oss-safeguard-20b
284
qwen1-5-moe-a2-7b
56.6K/mo
7.9/mo
14.3B
2024-02-29 2yr 5mo ago
Qwen/Qwen1.5-MoE-A2.7B
285
olmoe-1b-7b-0924
55.4K/mo
6/mo
6.92B
2024-07-20 2yr ago
allenai/OLMoE-1B-7B-0924
286
qwen-72b
55.2K/mo
11.2/mo
72.3B
2023-11-26 2yr 8mo ago
Qwen/Qwen-72B
287
nemotron-mini-4b-instruct
52.4K/mo
8.2/mo
unknown
2024-09-10 1yr 10mo ago
nvidia/Nemotron-Mini-4B-Instruct
288
llama-3-70b-instruct
49.4K/mo
2.6/mo
70.6B
2024-04-18 2yr 3mo ago
casperhansen/llama-3-70b-instruct-awq
289
ternary-bonsai-8b
49.3K/mo
37.8/mo
8.19B
2026-04-18 3mo ago
prism-ml/Ternary-Bonsai-8B-gguf
290
gemma-4-e4b-agentic-opus-reasoning-geminicli
47.8K/mo
8.1/mo
unknown
2026-04-05 3mo ago
deadbydawn101/gemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-mlx-4bit
291
llama-guard-4-12b-quantized-w4a16
45K/mo
0/mo
4.14B
2026-02-16 5mo ago
RedHatAI/Llama-Guard-4-12B-quantized.w4a16
292
qwen3guard-gen-4b
43.4K/mo
5.1/mo
4.41B
2025-09-23 10mo ago
Qwen/Qwen3Guard-Gen-4B
293
apertus-70b-instruct-2509-quantized-w4a16
42.5K/mo
<0.1/mo
11.3B
2025-09-21 10mo ago
RedHatAI/Apertus-70B-Instruct-2509-quantized.w4a16
294
olmoe-1b-7b-0125-instruct
41.5K/mo
3.8/mo
6.92B
2025-01-27 1yr 6mo ago
allenai/OLMoE-1B-7B-0125-Instruct
295
wildguard
41.1K/mo
2.2/mo
7.25B
2024-06-15 2yr 1mo ago
allenai/wildguard
296
qwen1-5-moe-a2-7b-chat-quantized-w4a16
38.8K/mo
<0.1/mo
14.4B
2025-02-24 1yr 5mo ago
nm-testing/Qwen1.5-MoE-A2.7B-Chat-quantized.w4a16
297
ui-tars-1-5-7b
38.7K/mo
1.8/mo
7.62B
2025-04-17 1yr 3mo ago
mradermacher/UI-TARS-1.5-7B-GGUF
298
llama-3-2-90b-vision-instruct
38.5K/mo
16.1/mo
88.6B
2024-09-19 1yr 10mo ago
meta-llama/Llama-3.2-90B-Vision-Instruct
299
mistral-7b-instruct-v0-3
36.9K/mo
0.4/mo
7.25B
2024-05-23 2yr 2mo ago
solidrust/Mistral-7B-Instruct-v0.3-AWQ
300
qwen3-coder-next-nvfp4-gb10
34K/mo
0.6/mo
unknown
2026-04-16 3mo ago
gdubicki/Qwen3-Coder-Next-NVFP4-GB10
301
internvl3-78b
32.7K/mo
0.6/mo
unknown
2025-04-17 1yr 3mo ago
OpenGVLab/InternVL3-78B-AWQ
302
nvidia-nemotron-nano-12b-v2-vl-nvfp4-qad
31.9K/mo
3/mo
7.70B
2025-10-22 9mo ago
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-NVFP4-QAD
303
t5gemma-2-270m-270m
31.1K/mo
21.8/mo
0.79B
2025-10-25 9mo ago
google/t5gemma-2-270m-270m
304
phi-4-multimodal-instruct
28K/mo
1.2/mo
4.17B
2025-09-05 10mo ago
nvidia/Phi-4-multimodal-instruct-NVFP4
305
paligemma-3b-ft-cococap-448
25.7K/mo
0.1/mo
2.92B
2024-05-13 2yr 2mo ago
google/paligemma-3b-ft-cococap-448
306
exaone-3-5-7-8b-instruct
24.6K/mo
0.9/mo
7.82B
2024-12-01 1yr 7mo ago
LGAI-EXAONE/EXAONE-3.5-7.8B-Instruct-AWQ
307
mistral-large-instruct-2411
19.9K/mo
0.4/mo
123B
2024-11-19 1yr 8mo ago
TechxGenus/Mistral-Large-Instruct-2411-AWQ
308
gemma-3-12b-it-abliterated-gptqmodel-4b-128g
17.2K/mo
<0.1/mo
12.2B
2025-04-14 1yr 3mo ago
jeffcookio/gemma-3-12b-it-abliterated-gptqmodel-4b-128g
309
ministral-8b-instruct-2410
16.5K/mo
<0.1/mo
8.02B
2024-10-24 1yr 9mo ago
PyrTools/Ministral-8B-Instruct-2410-AWQ
310
aya-expanse-8b
9.7K/mo
<0.1/mo
9.08B
2024-10-26 1yr 9mo ago
Orion-zhen/aya-expanse-8b-AWQ
311
qwen3-4b-instruct-2507-nvfp4a16
8.7K/mo
<0.1/mo
2.43B
2025-08-07 11mo ago
apolloparty/Qwen3-4B-Instruct-2507-NVFP4A16