-
Qwen/Qwen3-Reranker-0.6B
Text Ranking • 0.6B • Updated • 2.58M • 379 -
jinaai/jina-reranker-m0
Text Classification • 2B • Updated • 215k • 120 -
jinaai/jina-reranker-v2-base-multilingual
Text Ranking • 0.3B • Updated • 975k • 353 -
jinaai/jina-embeddings-v2-base-en
Feature Extraction • 0.1B • Updated • 361k • 735
Bjorn Melin
BjornMelin
AI & ML interests
Large Language Models, AI Agents, Multi-Agent Orchestrations, Deep Learning, NLP, Local LLM Optimization.
Recent Activity
liked a model about 2 months ago
Hcompany/Holo-3.1-4B updated a collection 4 months ago
Google liked a model 4 months ago
unsloth/gemma-4-26B-A4B-it-GGUFOrganizations
None yet
Datasets
Fine Tuning
- Running79
GGUF Model VRAM Calculator
📈79Calculate VRAM requirements for LLM models
- Running on CPU UpgradeAgentsFeatured1.01k
Model Memory Utility
🚀1.01kCalculate GPU memory needed for training Hugging Face models
- RunningFeatured1.05k
Can You Run It? LLM version
🚀1.05kCheck if your GPU can run a chosen LLM model
Legendary VL Models
Smol Models
My favorite smaller models under 10B parameters.
-
unsloth/DeepSeek-R1-0528-Qwen3-8B-GGUF
Text Generation • 8B • Updated • 52.8k • 436 -
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Text Generation • 8B • Updated • 52.5k • • 221 -
deepseek-ai/DeepSeek-R1-Distill-Llama-8B
Text Generation • 8B • Updated • 409k • • 872 -
Qwen/Qwen2.5-Coder-7B-Instruct
Text Generation • 8B • Updated • 1.91M • • 763
Llama
-
MaziyarPanahi/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 62.9k • 15 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.76M • • 2.37k -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 8.02M • • 6.41k -
MaziyarPanahi/Meta-Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 69.1k • 37
LLMs
-
deepseek-ai/DeepSeek-V3
Text Generation • 685B • Updated • 1.16M • • 4.1k -
sentence-transformers/static-retrieval-mrl-en-v1
Sentence Similarity • Updated • 59 -
internlm/internlm3-8b-instruct
Text Generation • 9B • Updated • 89k • 232 -
NovaSky-AI/Sky-T1-32B-Preview
Text Generation • 33B • Updated • 31 • • 550
Embedding Models
Single 4090 Laptop GPU
-
nvidia/OpenReasoning-Nemotron-32B
Text Generation • 33B • Updated • 9.43k • • 125 -
Qwen/Qwen3-32B-AWQ
Text Generation • 33B • Updated • 2.01M • 136 -
OpenHands/openhands-lm-32b-v0.1
Text Generation • 33B • Updated • 129 • 392 -
deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
Text Generation • 15B • Updated • 504k • • 666
Leaderboards
- Runtime errorFeatured142
smolagents LLM leaderboard
🏆142A leaderboard for LLMs powering smolagents
- RunningFeatured476
LLM Performance Leaderboard
🐨476View the LLM leaderboard rankings
- RunningAgentsFeatured222
Low-bit LLM Leaderboard
🏆222Track, rank and evaluate open LLMs and chatbots
- Running1.94k
UGI Leaderboard
📢1.94kUncensored General Intelligence Leaderboard
Coding Models
Google
-
google/gemma-3-27b-it-qat-q4_0-gguf
Image-Text-to-Text • 27B • Updated • 237 • 400 -
unsloth/gemma-3-27b-it-GGUF
Image-Text-to-Text • 27B • Updated • 13.6k • 207 -
google/gemma-3-27b-it
Image-Text-to-Text • 27B • Updated • 851k • • 2k -
google/gemma-3n-E4B-it
Image-Text-to-Text • 8B • Updated • 30.3k • • 919
Qwen
-
Qwen/Qwen3-30B-A3B-Instruct-2507
Text Generation • 31B • Updated • 2.12M • • 822 -
Qwen/Qwen3-235B-A22B-Thinking-2507
Text Generation • 235B • Updated • 20.4k • • 408 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 10.1M • • 722 -
unsloth/Qwen3-30B-A3B-GGUF
Text Generation • 31B • Updated • 56.5k • 284
Rerankers
-
Qwen/Qwen3-Reranker-0.6B
Text Ranking • 0.6B • Updated • 2.58M • 379 -
jinaai/jina-reranker-m0
Text Classification • 2B • Updated • 215k • 120 -
jinaai/jina-reranker-v2-base-multilingual
Text Ranking • 0.3B • Updated • 975k • 353 -
jinaai/jina-embeddings-v2-base-en
Feature Extraction • 0.1B • Updated • 361k • 735
Embedding Models
Datasets
Single 4090 Laptop GPU
-
nvidia/OpenReasoning-Nemotron-32B
Text Generation • 33B • Updated • 9.43k • • 125 -
Qwen/Qwen3-32B-AWQ
Text Generation • 33B • Updated • 2.01M • 136 -
OpenHands/openhands-lm-32b-v0.1
Text Generation • 33B • Updated • 129 • 392 -
deepseek-ai/DeepSeek-R1-Distill-Qwen-14B
Text Generation • 15B • Updated • 504k • • 666
Fine Tuning
- Running79
GGUF Model VRAM Calculator
📈79Calculate VRAM requirements for LLM models
- Running on CPU UpgradeAgentsFeatured1.01k
Model Memory Utility
🚀1.01kCalculate GPU memory needed for training Hugging Face models
- RunningFeatured1.05k
Can You Run It? LLM version
🚀1.05kCheck if your GPU can run a chosen LLM model
Leaderboards
- Runtime errorFeatured142
smolagents LLM leaderboard
🏆142A leaderboard for LLMs powering smolagents
- RunningFeatured476
LLM Performance Leaderboard
🐨476View the LLM leaderboard rankings
- RunningAgentsFeatured222
Low-bit LLM Leaderboard
🏆222Track, rank and evaluate open LLMs and chatbots
- Running1.94k
UGI Leaderboard
📢1.94kUncensored General Intelligence Leaderboard
Legendary VL Models
Coding Models
Smol Models
My favorite smaller models under 10B parameters.
-
unsloth/DeepSeek-R1-0528-Qwen3-8B-GGUF
Text Generation • 8B • Updated • 52.8k • 436 -
nvidia/Llama-3.1-Nemotron-Nano-8B-v1
Text Generation • 8B • Updated • 52.5k • • 221 -
deepseek-ai/DeepSeek-R1-Distill-Llama-8B
Text Generation • 8B • Updated • 409k • • 872 -
Qwen/Qwen2.5-Coder-7B-Instruct
Text Generation • 8B • Updated • 1.91M • • 763
Google
-
google/gemma-3-27b-it-qat-q4_0-gguf
Image-Text-to-Text • 27B • Updated • 237 • 400 -
unsloth/gemma-3-27b-it-GGUF
Image-Text-to-Text • 27B • Updated • 13.6k • 207 -
google/gemma-3-27b-it
Image-Text-to-Text • 27B • Updated • 851k • • 2k -
google/gemma-3n-E4B-it
Image-Text-to-Text • 8B • Updated • 30.3k • • 919
Llama
-
MaziyarPanahi/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 62.9k • 15 -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.76M • • 2.37k -
meta-llama/Llama-3.1-8B-Instruct
Text Generation • 8B • Updated • 8.02M • • 6.41k -
MaziyarPanahi/Meta-Llama-3.1-8B-Instruct-GGUF
Text Generation • 8B • Updated • 69.1k • 37
Qwen
-
Qwen/Qwen3-30B-A3B-Instruct-2507
Text Generation • 31B • Updated • 2.12M • • 822 -
Qwen/Qwen3-235B-A22B-Thinking-2507
Text Generation • 235B • Updated • 20.4k • • 408 -
Qwen/Qwen3-32B
Text Generation • 33B • Updated • 10.1M • • 722 -
unsloth/Qwen3-30B-A3B-GGUF
Text Generation • 31B • Updated • 56.5k • 284
LLMs
-
deepseek-ai/DeepSeek-V3
Text Generation • 685B • Updated • 1.16M • • 4.1k -
sentence-transformers/static-retrieval-mrl-en-v1
Sentence Similarity • Updated • 59 -
internlm/internlm3-8b-instruct
Text Generation • 9B • Updated • 89k • 232 -
NovaSky-AI/Sky-T1-32B-Preview
Text Generation • 33B • Updated • 31 • • 550