New 2.0 version of our Dynamic GGUF + Quants. Dynamic 2.0 achieves superior accuracy & SOTA quantization performance.
AI & ML interests
Open Source AI 🦥
Recent Activity
View all activity
Gemma 4 is Google's new model family including including E2B, E4B, 26B-A4B, and 31B.
-
unsloth/gemma-4-12b-it-GGUF
Image-Text-to-Text • 12B • Updated • 627k • 777 -
unsloth/gemma-4-26B-A4B-it-GGUF
Image-Text-to-Text • 25B • Updated • 1.46M • 1.01k -
unsloth/gemma-4-31B-it-GGUF
Image-Text-to-Text • 31B • Updated • 488k • 559 -
unsloth/gemma-4-E4B-it-GGUF
Image-Text-to-Text • 8B • Updated • 513k • 570
Gemma 4 QAT (Quantization-Aware Training) for 3x less memory use and near original accuracy.
-
unsloth/gemma-4-12B-it-qat-GGUF
Any-to-Any • 12B • Updated • 246k • 380 -
unsloth/gemma-4-26B-A4B-it-qat-GGUF
Image-Text-to-Text • 25B • Updated • 417k • 349 -
unsloth/gemma-4-31B-it-qat-GGUF
Image-Text-to-Text • 31B • Updated • 198k • 169 -
unsloth/gemma-4-E4B-it-qat-GGUF
Any-to-Any • 7B • Updated • 280k • 147
Qwen3.5 is Qwen's new model family including Qwen3.5 Small: 0.8B, 2B, 4B, 9B and Qwen3.5 Medium: 35B-A3B, 27B, 122B-A10B and 397B-A17B.
-
unsloth/Qwen3.5-27B-GGUF
Image-Text-to-Text • 27B • Updated • 101k • 502 -
unsloth/Qwen3.5-35B-A3B-GGUF
Image-Text-to-Text • 35B • Updated • 155k • 863 -
unsloth/Qwen3.5-9B-GGUF
Image-Text-to-Text • 9B • Updated • 985k • 799 -
unsloth/Qwen3.5-4B-GGUF
Image-Text-to-Text • 4B • Updated • 1.17M • 353
The Qwen3-Coder models deliver SOTA advancements in agentic coding and code tasks. Includes Qwen3-Coder-Next.
-
unsloth/Qwen3-Coder-Next-GGUF
Text Generation • 80B • Updated • 202k • 788 -
unsloth/Qwen3-Coder-Next-FP8-Dynamic
Text Generation • 80B • Updated • 15.6k • 45 -
unsloth/Qwen3-Coder-Next
Text Generation • 80B • Updated • 982 • 27 -
unsloth/Qwen3-Coder-Next-FP8
Text Generation • 80B • Updated • 1.7k • 11
Qwen's new Qwen3 models. In Unsloth Dynamic 2.0, GGUF, 4-bit and 16-bit Safetensor formats. Includes 128K Context Length variants.
DeepSeek-R1-0528 is here! The most powerful reasoning open LLM, available in GGUF, original & 4-bit formats. Includes Llama & Qwen distilled models.
-
unsloth/DeepSeek-R1-0528-Qwen3-8B-GGUF
Text Generation • 8B • Updated • 50.9k • 437 -
unsloth/DeepSeek-R1-0528-GGUF
Text Generation • 671B • Updated • 7.55k • 202 -
unsloth/DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 4.46k • 13 -
unsloth/DeepSeek-R1-0528
Text Generation • 685B • Updated • 21 • 15
All versions of Google's new multimodal models including QAT in 1B, 4B, 12B, and 27B sizes. In GGUF, dynamic 4-bit and 16-bit formats.
-
unsloth/gemma-3-270m-it-GGUF
Text Generation • 0.3B • Updated • 83.9k • 169 -
unsloth/gemma-3-270m-it-qat-GGUF
Text Generation • 0.3B • Updated • 2.99k • 13 -
unsloth/gemma-3-270m-it
Text Generation • 0.3B • Updated • 22.6k • 23 -
unsloth/gemma-3-270m-it-unsloth-bnb-4bit
Text Generation • 0.3B • Updated • 6.62k • 6
IBM's new Granite-4.0 models! Run Dynamic GGUFs or fine-tune with Unsloth.
Meta's new Llama 4 multimodal models, Scout & Maverick. Includes Dynamic GGUFs, 16-bit & Dynamic 4-bit uploads. Run & fine-tune them with Unsloth!
-
unsloth/Llama-4-Scout-17B-16E-Instruct-GGUF
Image-Text-to-Text • 108B • Updated • 29.3k • 160 -
unsloth/Llama-4-Maverick-17B-128E-Instruct-GGUF
Image-Text-to-Text • 401B • Updated • 6.11k • 49 -
unsloth/Llama-4-Scout-17B-16E-Instruct
Image-Text-to-Text • 109B • Updated • 4.86k • 57 -
unsloth/Llama-4-Scout-17B-16E-Instruct-unsloth-bnb-4bit
Image-Text-to-Text • 112B • Updated • 187 • 80
Unsloths Dynamic 4bit Quants selectively skips quantizing certain parameters; greatly improving accuracy while only using <10% more VRAM than BnB 4bit
-
unsloth/DeepSeek-R1-Distill-Llama-8B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 13.3k • 37 -
unsloth/DeepSeek-R1-Distill-Qwen-14B-unsloth-bnb-4bit
Text Generation • 15B • Updated • 5.67k • 30 -
unsloth/DeepSeek-R1-Distill-Qwen-7B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 5.4k • 25 -
unsloth/gemma-3-12b-it-unsloth-bnb-4bit
Image-Text-to-Text • 12B • Updated • 29.5k • 24
A collection of 4-bit, Dynamic 4-bit and 16-bit voice models including Sesame-CSM, OpenAI's Whisper, Orpheus. Fine-tune them with Unsloth now!
-
unsloth/orpheus-3b-0.1-ft-GGUF
Text-to-Speech • 3B • Updated • 2.45k • 21 -
unsloth/orpheus-3b-0.1-ft-unsloth-bnb-4bit
Text-to-Speech • 3B • Updated • 3.61k • 19 -
unsloth/csm-1b
Text-to-Speech • 2B • Updated • 1.95k • 21 -
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4.63k • 21
Meta's new Llama 3.2 vision and text models including 1B, 3B, 11B and 90B. Includes GGUF, 4-bit bnb and original versions.
-
unsloth/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 18.4k • 69 -
unsloth/Llama-3.2-1B-Instruct
Text Generation • 1B • Updated • 301k • 102 -
unsloth/Llama-3.2-1B-Instruct-unsloth-bnb-4bit
Text Generation • 1B • Updated • 64.2k • 4 -
unsloth/Llama-3.2-1B-Instruct-bnb-4bit
Text Generation • 1B • Updated • 40.8k • 23
All versions of Qwen2.5-VL including the new 32B version and 4-bit, 16-bit and more!
-
unsloth/Qwen2.5-VL-3B-Instruct-GGUF
Image-Text-to-Text • 3B • Updated • 7.64k • 27 -
unsloth/Qwen2.5-VL-7B-Instruct-GGUF
Image-Text-to-Text • 8B • Updated • 178k • 200 -
unsloth/Qwen2.5-VL-32B-Instruct-GGUF
Image-Text-to-Text • 33B • Updated • 3.1k • 9 -
unsloth/Qwen2.5-VL-72B-Instruct-GGUF
Image-Text-to-Text • 73B • Updated • 4.56k • 11
Collection of the most popular vision models including Llama 3.2, LlaVa, Qwen2 VL, Pixtral, PaliGemma and more!
-
unsloth/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 8.04k • 89 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Llama-3.2-11B-Vision-Instruct-unsloth-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.23k • 29 -
unsloth/Qwen2-VL-7B-Instruct-bnb-4bit
Image-Text-to-Text • 9B • Updated • 1.62k • 6
Complete collection of Code-specific model series for Qwen2.5 in bnb 4bit, 16bit and GGUF formats.
-
unsloth/Qwen2.5-Coder-32B-Instruct-128K-GGUF
33B • Updated • 5.24k • 79 -
unsloth/Qwen2.5-Coder-14B-Instruct-128K-GGUF
15B • Updated • 12.3k • 47 -
unsloth/Qwen2.5-Coder-7B-Instruct-128K-GGUF
8B • Updated • 11.6k • 37 -
unsloth/Qwen2.5-Coder-3B-Instruct-128K-GGUF
3B • Updated • 5.19k • 27
-
unsloth/Qwen2.5-7B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 203k • 23 -
unsloth/Qwen2.5-7B-Instruct
Text Generation • 8B • Updated • 80k • • 27 -
unsloth/Qwen2.5-14B-bnb-4bit
Text Generation • 15B • Updated • 232k • 5 -
unsloth/Qwen2.5-7B-bnb-4bit
Text Generation • 8B • Updated • 10.5k • 7
-
unsloth/Llama-3.2-3B-Instruct-bnb-4bit
Text Generation • 3B • Updated • 57.1k • 36 -
unsloth/Llama-3.2-1B-Instruct-bnb-4bit
Text Generation • 1B • Updated • 40.8k • 23 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 68.2k • 102
Dynamic Unsloth NVFP4 Quants. Faster and more accurate.
-
unsloth/gemma-4-31B-it-NVFP4
Image-Text-to-Text • 23B • Updated • 27.1k • 28 -
unsloth/gemma-4-26B-A4B-it-NVFP4
Image-Text-to-Text • 16B • Updated • 118k • 20 -
unsloth/gemma-4-12b-it-NVFP4
Image-Text-to-Text • 8B • Updated • 37.8k • 19 -
unsloth/gemma-4-E4B-it-NVFP4
Image-Text-to-Text • 7B • Updated • 10.6k • 10
-
unsloth/Qwen3.6-27B-NVFP4
Image-Text-to-Text • 21B • Updated • 2.91M • 260 -
unsloth/Qwen3.6-35B-A3B-NVFP4
Image-Text-to-Text • 25B • Updated • 1.16M • 104 -
unsloth/Qwen3.6-35B-A3B-NVFP4-Fast
Image-Text-to-Text • 22B • Updated • 363k • 92 -
unsloth/Qwen3.6-27B-GGUF
Image-Text-to-Text • 27B • Updated • 717k • 915
Find GGUFs and other variants of diffusion based models like Qwen-Image and FLUX.
-
unsloth/LTX-2.3-GGUF
Image-to-Video • 21B • Updated • 271k • 536 -
unsloth/Qwen-Image-Edit-2511-GGUF
Image-to-Image • 20B • Updated • 199k • 550 -
unsloth/Qwen-Image-2512-GGUF
Text-to-Image • 20B • Updated • 80.1k • 392 -
unsloth/LTX-2-GGUF
Image-to-Video • 19B • Updated • 4.58k • 133
OpenAI's gpt-oss-20b and gpt-oss-120b is here! The powerful open models are available in GGUF, original & 4-bit formats.
-
unsloth/gpt-oss-20b-GGUF
Text Generation • 21B • Updated • 557k • 753 -
unsloth/gpt-oss-120b-GGUF
Text Generation • 117B • Updated • 108k • 287 -
unsloth/gpt-oss-20b-unsloth-bnb-4bit
Text Generation • 21B • Updated • 65.3k • 38 -
unsloth/gpt-oss-120b-unsloth-bnb-4bit
Text Generation • 117B • Updated • 9.85k • 17
Qwen's new multimodal vision models in GGUF, safetensor, and dynamic Unsloth formats.
-
unsloth/Qwen3-VL-30B-A3B-Instruct-GGUF
Image-Text-to-Text • 31B • Updated • 74.6k • 106 -
unsloth/Qwen3-VL-30B-A3B-Thinking-GGUF
Image-Text-to-Text • 31B • Updated • 4.39k • 41 -
unsloth/Qwen3-VL-4B-Instruct-GGUF
Image-Text-to-Text • 4B • Updated • 241k • 63 -
unsloth/Qwen3-VL-4B-Thinking-GGUF
Image-Text-to-Text • 4B • Updated • 3.59k • 25
DeepSeek's new 3.1 update to their V3 models!
Run or fine-tune embedding models with Unsloth.
-
unsloth/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 21.6k • • 9 -
unsloth/embeddinggemma-300m-GGUF
Sentence Similarity • 0.3B • Updated • 10.1k • 82 -
unsloth/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 5.15k • 9 -
unsloth/Qwen3-Embedding-4B
Feature Extraction • 4B • Updated • 313k • 4
Mistral Ministral 3: new multimodal models in Base, Instruct, and Reasoning variants, available in 3B, 8B, and 14B sizes.
Google Gemma 3n models, all versions including Dynamic GGUF, 4-bit, 16-bit and formats!
-
unsloth/gemma-3n-E4B-it-GGUF
Image-Text-to-Text • 7B • Updated • 6.06k • 206 -
unsloth/gemma-3n-E2B-it-GGUF
Image-Text-to-Text • 4B • Updated • 11.8k • 60 -
unsloth/gemma-3n-E4B-it-unsloth-bnb-4bit
Image-Text-to-Text • 8B • Updated • 2.24k • 9 -
unsloth/gemma-3n-E4B-unsloth-bnb-4bit
Image-Text-to-Text • 8B • Updated • 310 • 4
Microsoft's Phi-4 models including Reasoning + Reasoning Plus & mini. Includes Dynamic 2.0 GGUF, 4-bit & 16-bit versions. Includes Unsloth's bug fixes
-
unsloth/Phi-4-reasoning-plus-GGUF
Text Generation • 15B • Updated • 8.93k • 95 -
unsloth/Phi-4-mini-reasoning-GGUF
Text Generation • 4B • Updated • 11k • 76 -
unsloth/Phi-4-reasoning-GGUF
Text Generation • 15B • Updated • 3.14k • 23 -
unsloth/phi-4-GGUF
Text Generation • 15B • Updated • 5.78k • 191
Deepseek-V3-0324 and V3 - available in original, and Dynamic GGUF formats, with support for 2-8-bit quantized versions.
-
unsloth/DeepSeek-V3-0324-GGUF-UD
Text Generation • 671B • Updated • 1.86k • 23 -
unsloth/DeepSeek-V3-0324-GGUF
Text Generation • 671B • Updated • 4.77k • 200 -
unsloth/DeepSeek-V3-0324
Text Generation • 684B • Updated • 11 • 8 -
unsloth/DeepSeek-V3-0324-BF16
Text Generation • 684B • Updated • 18 • 4
A collection of Mistral's new Small 3.2 and 3 models including GGUF, 4-bit and more!
-
unsloth/Mistral-Small-3.2-24B-Instruct-2506-GGUF
Image-Text-to-Text • 24B • Updated • 39.9k • 185 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506
Image-Text-to-Text • 24B • Updated • 7.61k • • 16 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506-FP8
Image-Text-to-Text • Updated • 74 • 6 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506-unsloth-bnb-4bit
Image-Text-to-Text • 25B • Updated • 2.51k • 13
Meta's new Llama 3.3 (70B) model in all formats. Includes GGUF, 4-bit bnb and original versions.
Qwen's reasoning models including QwQ (32B) & QVQ (72B) in formats: GGUF, dynamic 4-bit and 16-bit original versions.
-
unsloth/QwQ-32B-GGUF
Text Generation • 33B • Updated • 3.72k • 86 -
unsloth/QwQ-32B-unsloth-bnb-4bit
Text Generation • 34B • Updated • 120 • 47 -
unsloth/QwQ-32B
Text Generation • 33B • Updated • 17 • • 17 -
unsloth/QwQ-32B-bnb-4bit
Text Generation • 34B • Updated • 33 • 4
Meta's Llama 3.2 vision models 11B and 90B. Include 4-bit bnb and original versions.
-
unsloth/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 8.04k • 89 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Llama-3.2-11B-Vision
Image-Text-to-Text • 11B • Updated • 241 • 35 -
unsloth/Llama-3.2-11B-Vision-bnb-4bit
Image-Text-to-Text • 11B • Updated • 45 • 16
Meta's Llama 3.1 models including 8B, 70B, 405B. Includes 4-bit bnb and original versions.
-
unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 68.2k • 102 -
unsloth/Llama-3.1-8B-Instruct-unsloth-bnb-4bit
Text Generation • 8B • Updated • 24.1k • 4 -
unsloth/Meta-Llama-3.1-8B-Instruct-unsloth-bnb-4bit
Text Generation • 8B • Updated • 39.2k • 4 -
unsloth/Meta-Llama-3.1-8B-bnb-4bit
Text Generation • 8B • Updated • 23.7k • 112
Native bitsandbytes 4bit pre quantized models
-
unsloth/Llama-3.2-3B-bnb-4bit
Text Generation • 3B • Updated • 11.3k • 23 -
unsloth/Meta-Llama-3.1-8B-bnb-4bit
Text Generation • 8B • Updated • 23.7k • 112 -
unsloth/llama-3-8b-Instruct-bnb-4bit
Text Generation • 8B • Updated • 47k • 134 -
unsloth/gemma-2-9b-bnb-4bit
Text Generation • 10B • Updated • 10.4k • 31
New 2.0 version of our Dynamic GGUF + Quants. Dynamic 2.0 achieves superior accuracy & SOTA quantization performance.
Dynamic Unsloth NVFP4 Quants. Faster and more accurate.
-
unsloth/gemma-4-31B-it-NVFP4
Image-Text-to-Text • 23B • Updated • 27.1k • 28 -
unsloth/gemma-4-26B-A4B-it-NVFP4
Image-Text-to-Text • 16B • Updated • 118k • 20 -
unsloth/gemma-4-12b-it-NVFP4
Image-Text-to-Text • 8B • Updated • 37.8k • 19 -
unsloth/gemma-4-E4B-it-NVFP4
Image-Text-to-Text • 7B • Updated • 10.6k • 10
Gemma 4 is Google's new model family including including E2B, E4B, 26B-A4B, and 31B.
-
unsloth/gemma-4-12b-it-GGUF
Image-Text-to-Text • 12B • Updated • 627k • 777 -
unsloth/gemma-4-26B-A4B-it-GGUF
Image-Text-to-Text • 25B • Updated • 1.46M • 1.01k -
unsloth/gemma-4-31B-it-GGUF
Image-Text-to-Text • 31B • Updated • 488k • 559 -
unsloth/gemma-4-E4B-it-GGUF
Image-Text-to-Text • 8B • Updated • 513k • 570
-
unsloth/Qwen3.6-27B-NVFP4
Image-Text-to-Text • 21B • Updated • 2.91M • 260 -
unsloth/Qwen3.6-35B-A3B-NVFP4
Image-Text-to-Text • 25B • Updated • 1.16M • 104 -
unsloth/Qwen3.6-35B-A3B-NVFP4-Fast
Image-Text-to-Text • 22B • Updated • 363k • 92 -
unsloth/Qwen3.6-27B-GGUF
Image-Text-to-Text • 27B • Updated • 717k • 915
Gemma 4 QAT (Quantization-Aware Training) for 3x less memory use and near original accuracy.
-
unsloth/gemma-4-12B-it-qat-GGUF
Any-to-Any • 12B • Updated • 246k • 380 -
unsloth/gemma-4-26B-A4B-it-qat-GGUF
Image-Text-to-Text • 25B • Updated • 417k • 349 -
unsloth/gemma-4-31B-it-qat-GGUF
Image-Text-to-Text • 31B • Updated • 198k • 169 -
unsloth/gemma-4-E4B-it-qat-GGUF
Any-to-Any • 7B • Updated • 280k • 147
Find GGUFs and other variants of diffusion based models like Qwen-Image and FLUX.
-
unsloth/LTX-2.3-GGUF
Image-to-Video • 21B • Updated • 271k • 536 -
unsloth/Qwen-Image-Edit-2511-GGUF
Image-to-Image • 20B • Updated • 199k • 550 -
unsloth/Qwen-Image-2512-GGUF
Text-to-Image • 20B • Updated • 80.1k • 392 -
unsloth/LTX-2-GGUF
Image-to-Video • 19B • Updated • 4.58k • 133
Qwen3.5 is Qwen's new model family including Qwen3.5 Small: 0.8B, 2B, 4B, 9B and Qwen3.5 Medium: 35B-A3B, 27B, 122B-A10B and 397B-A17B.
-
unsloth/Qwen3.5-27B-GGUF
Image-Text-to-Text • 27B • Updated • 101k • 502 -
unsloth/Qwen3.5-35B-A3B-GGUF
Image-Text-to-Text • 35B • Updated • 155k • 863 -
unsloth/Qwen3.5-9B-GGUF
Image-Text-to-Text • 9B • Updated • 985k • 799 -
unsloth/Qwen3.5-4B-GGUF
Image-Text-to-Text • 4B • Updated • 1.17M • 353
OpenAI's gpt-oss-20b and gpt-oss-120b is here! The powerful open models are available in GGUF, original & 4-bit formats.
-
unsloth/gpt-oss-20b-GGUF
Text Generation • 21B • Updated • 557k • 753 -
unsloth/gpt-oss-120b-GGUF
Text Generation • 117B • Updated • 108k • 287 -
unsloth/gpt-oss-20b-unsloth-bnb-4bit
Text Generation • 21B • Updated • 65.3k • 38 -
unsloth/gpt-oss-120b-unsloth-bnb-4bit
Text Generation • 117B • Updated • 9.85k • 17
The Qwen3-Coder models deliver SOTA advancements in agentic coding and code tasks. Includes Qwen3-Coder-Next.
-
unsloth/Qwen3-Coder-Next-GGUF
Text Generation • 80B • Updated • 202k • 788 -
unsloth/Qwen3-Coder-Next-FP8-Dynamic
Text Generation • 80B • Updated • 15.6k • 45 -
unsloth/Qwen3-Coder-Next
Text Generation • 80B • Updated • 982 • 27 -
unsloth/Qwen3-Coder-Next-FP8
Text Generation • 80B • Updated • 1.7k • 11
Qwen's new multimodal vision models in GGUF, safetensor, and dynamic Unsloth formats.
-
unsloth/Qwen3-VL-30B-A3B-Instruct-GGUF
Image-Text-to-Text • 31B • Updated • 74.6k • 106 -
unsloth/Qwen3-VL-30B-A3B-Thinking-GGUF
Image-Text-to-Text • 31B • Updated • 4.39k • 41 -
unsloth/Qwen3-VL-4B-Instruct-GGUF
Image-Text-to-Text • 4B • Updated • 241k • 63 -
unsloth/Qwen3-VL-4B-Thinking-GGUF
Image-Text-to-Text • 4B • Updated • 3.59k • 25
Qwen's new Qwen3 models. In Unsloth Dynamic 2.0, GGUF, 4-bit and 16-bit Safetensor formats. Includes 128K Context Length variants.
DeepSeek's new 3.1 update to their V3 models!
DeepSeek-R1-0528 is here! The most powerful reasoning open LLM, available in GGUF, original & 4-bit formats. Includes Llama & Qwen distilled models.
-
unsloth/DeepSeek-R1-0528-Qwen3-8B-GGUF
Text Generation • 8B • Updated • 50.9k • 437 -
unsloth/DeepSeek-R1-0528-GGUF
Text Generation • 671B • Updated • 7.55k • 202 -
unsloth/DeepSeek-R1-0528-Qwen3-8B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 4.46k • 13 -
unsloth/DeepSeek-R1-0528
Text Generation • 685B • Updated • 21 • 15
Run or fine-tune embedding models with Unsloth.
-
unsloth/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 21.6k • • 9 -
unsloth/embeddinggemma-300m-GGUF
Sentence Similarity • 0.3B • Updated • 10.1k • 82 -
unsloth/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 5.15k • 9 -
unsloth/Qwen3-Embedding-4B
Feature Extraction • 4B • Updated • 313k • 4
All versions of Google's new multimodal models including QAT in 1B, 4B, 12B, and 27B sizes. In GGUF, dynamic 4-bit and 16-bit formats.
-
unsloth/gemma-3-270m-it-GGUF
Text Generation • 0.3B • Updated • 83.9k • 169 -
unsloth/gemma-3-270m-it-qat-GGUF
Text Generation • 0.3B • Updated • 2.99k • 13 -
unsloth/gemma-3-270m-it
Text Generation • 0.3B • Updated • 22.6k • 23 -
unsloth/gemma-3-270m-it-unsloth-bnb-4bit
Text Generation • 0.3B • Updated • 6.62k • 6
Mistral Ministral 3: new multimodal models in Base, Instruct, and Reasoning variants, available in 3B, 8B, and 14B sizes.
IBM's new Granite-4.0 models! Run Dynamic GGUFs or fine-tune with Unsloth.
Google Gemma 3n models, all versions including Dynamic GGUF, 4-bit, 16-bit and formats!
-
unsloth/gemma-3n-E4B-it-GGUF
Image-Text-to-Text • 7B • Updated • 6.06k • 206 -
unsloth/gemma-3n-E2B-it-GGUF
Image-Text-to-Text • 4B • Updated • 11.8k • 60 -
unsloth/gemma-3n-E4B-it-unsloth-bnb-4bit
Image-Text-to-Text • 8B • Updated • 2.24k • 9 -
unsloth/gemma-3n-E4B-unsloth-bnb-4bit
Image-Text-to-Text • 8B • Updated • 310 • 4
Meta's new Llama 4 multimodal models, Scout & Maverick. Includes Dynamic GGUFs, 16-bit & Dynamic 4-bit uploads. Run & fine-tune them with Unsloth!
-
unsloth/Llama-4-Scout-17B-16E-Instruct-GGUF
Image-Text-to-Text • 108B • Updated • 29.3k • 160 -
unsloth/Llama-4-Maverick-17B-128E-Instruct-GGUF
Image-Text-to-Text • 401B • Updated • 6.11k • 49 -
unsloth/Llama-4-Scout-17B-16E-Instruct
Image-Text-to-Text • 109B • Updated • 4.86k • 57 -
unsloth/Llama-4-Scout-17B-16E-Instruct-unsloth-bnb-4bit
Image-Text-to-Text • 112B • Updated • 187 • 80
Microsoft's Phi-4 models including Reasoning + Reasoning Plus & mini. Includes Dynamic 2.0 GGUF, 4-bit & 16-bit versions. Includes Unsloth's bug fixes
-
unsloth/Phi-4-reasoning-plus-GGUF
Text Generation • 15B • Updated • 8.93k • 95 -
unsloth/Phi-4-mini-reasoning-GGUF
Text Generation • 4B • Updated • 11k • 76 -
unsloth/Phi-4-reasoning-GGUF
Text Generation • 15B • Updated • 3.14k • 23 -
unsloth/phi-4-GGUF
Text Generation • 15B • Updated • 5.78k • 191
Unsloths Dynamic 4bit Quants selectively skips quantizing certain parameters; greatly improving accuracy while only using <10% more VRAM than BnB 4bit
-
unsloth/DeepSeek-R1-Distill-Llama-8B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 13.3k • 37 -
unsloth/DeepSeek-R1-Distill-Qwen-14B-unsloth-bnb-4bit
Text Generation • 15B • Updated • 5.67k • 30 -
unsloth/DeepSeek-R1-Distill-Qwen-7B-unsloth-bnb-4bit
Text Generation • 8B • Updated • 5.4k • 25 -
unsloth/gemma-3-12b-it-unsloth-bnb-4bit
Image-Text-to-Text • 12B • Updated • 29.5k • 24
Deepseek-V3-0324 and V3 - available in original, and Dynamic GGUF formats, with support for 2-8-bit quantized versions.
-
unsloth/DeepSeek-V3-0324-GGUF-UD
Text Generation • 671B • Updated • 1.86k • 23 -
unsloth/DeepSeek-V3-0324-GGUF
Text Generation • 671B • Updated • 4.77k • 200 -
unsloth/DeepSeek-V3-0324
Text Generation • 684B • Updated • 11 • 8 -
unsloth/DeepSeek-V3-0324-BF16
Text Generation • 684B • Updated • 18 • 4
A collection of 4-bit, Dynamic 4-bit and 16-bit voice models including Sesame-CSM, OpenAI's Whisper, Orpheus. Fine-tune them with Unsloth now!
-
unsloth/orpheus-3b-0.1-ft-GGUF
Text-to-Speech • 3B • Updated • 2.45k • 21 -
unsloth/orpheus-3b-0.1-ft-unsloth-bnb-4bit
Text-to-Speech • 3B • Updated • 3.61k • 19 -
unsloth/csm-1b
Text-to-Speech • 2B • Updated • 1.95k • 21 -
unsloth/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4.63k • 21
A collection of Mistral's new Small 3.2 and 3 models including GGUF, 4-bit and more!
-
unsloth/Mistral-Small-3.2-24B-Instruct-2506-GGUF
Image-Text-to-Text • 24B • Updated • 39.9k • 185 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506
Image-Text-to-Text • 24B • Updated • 7.61k • • 16 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506-FP8
Image-Text-to-Text • Updated • 74 • 6 -
unsloth/Mistral-Small-3.2-24B-Instruct-2506-unsloth-bnb-4bit
Image-Text-to-Text • 25B • Updated • 2.51k • 13
Meta's new Llama 3.2 vision and text models including 1B, 3B, 11B and 90B. Includes GGUF, 4-bit bnb and original versions.
-
unsloth/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 18.4k • 69 -
unsloth/Llama-3.2-1B-Instruct
Text Generation • 1B • Updated • 301k • 102 -
unsloth/Llama-3.2-1B-Instruct-unsloth-bnb-4bit
Text Generation • 1B • Updated • 64.2k • 4 -
unsloth/Llama-3.2-1B-Instruct-bnb-4bit
Text Generation • 1B • Updated • 40.8k • 23
Meta's new Llama 3.3 (70B) model in all formats. Includes GGUF, 4-bit bnb and original versions.
All versions of Qwen2.5-VL including the new 32B version and 4-bit, 16-bit and more!
-
unsloth/Qwen2.5-VL-3B-Instruct-GGUF
Image-Text-to-Text • 3B • Updated • 7.64k • 27 -
unsloth/Qwen2.5-VL-7B-Instruct-GGUF
Image-Text-to-Text • 8B • Updated • 178k • 200 -
unsloth/Qwen2.5-VL-32B-Instruct-GGUF
Image-Text-to-Text • 33B • Updated • 3.1k • 9 -
unsloth/Qwen2.5-VL-72B-Instruct-GGUF
Image-Text-to-Text • 73B • Updated • 4.56k • 11
Qwen's reasoning models including QwQ (32B) & QVQ (72B) in formats: GGUF, dynamic 4-bit and 16-bit original versions.
-
unsloth/QwQ-32B-GGUF
Text Generation • 33B • Updated • 3.72k • 86 -
unsloth/QwQ-32B-unsloth-bnb-4bit
Text Generation • 34B • Updated • 120 • 47 -
unsloth/QwQ-32B
Text Generation • 33B • Updated • 17 • • 17 -
unsloth/QwQ-32B-bnb-4bit
Text Generation • 34B • Updated • 33 • 4
Collection of the most popular vision models including Llama 3.2, LlaVa, Qwen2 VL, Pixtral, PaliGemma and more!
-
unsloth/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 8.04k • 89 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Llama-3.2-11B-Vision-Instruct-unsloth-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.23k • 29 -
unsloth/Qwen2-VL-7B-Instruct-bnb-4bit
Image-Text-to-Text • 9B • Updated • 1.62k • 6
Meta's Llama 3.2 vision models 11B and 90B. Include 4-bit bnb and original versions.
-
unsloth/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 8.04k • 89 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Llama-3.2-11B-Vision
Image-Text-to-Text • 11B • Updated • 241 • 35 -
unsloth/Llama-3.2-11B-Vision-bnb-4bit
Image-Text-to-Text • 11B • Updated • 45 • 16
Complete collection of Code-specific model series for Qwen2.5 in bnb 4bit, 16bit and GGUF formats.
-
unsloth/Qwen2.5-Coder-32B-Instruct-128K-GGUF
33B • Updated • 5.24k • 79 -
unsloth/Qwen2.5-Coder-14B-Instruct-128K-GGUF
15B • Updated • 12.3k • 47 -
unsloth/Qwen2.5-Coder-7B-Instruct-128K-GGUF
8B • Updated • 11.6k • 37 -
unsloth/Qwen2.5-Coder-3B-Instruct-128K-GGUF
3B • Updated • 5.19k • 27
Meta's Llama 3.1 models including 8B, 70B, 405B. Includes 4-bit bnb and original versions.
-
unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 68.2k • 102 -
unsloth/Llama-3.1-8B-Instruct-unsloth-bnb-4bit
Text Generation • 8B • Updated • 24.1k • 4 -
unsloth/Meta-Llama-3.1-8B-Instruct-unsloth-bnb-4bit
Text Generation • 8B • Updated • 39.2k • 4 -
unsloth/Meta-Llama-3.1-8B-bnb-4bit
Text Generation • 8B • Updated • 23.7k • 112
-
unsloth/Qwen2.5-7B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 203k • 23 -
unsloth/Qwen2.5-7B-Instruct
Text Generation • 8B • Updated • 80k • • 27 -
unsloth/Qwen2.5-14B-bnb-4bit
Text Generation • 15B • Updated • 232k • 5 -
unsloth/Qwen2.5-7B-bnb-4bit
Text Generation • 8B • Updated • 10.5k • 7
Native bitsandbytes 4bit pre quantized models
-
unsloth/Llama-3.2-3B-bnb-4bit
Text Generation • 3B • Updated • 11.3k • 23 -
unsloth/Meta-Llama-3.1-8B-bnb-4bit
Text Generation • 8B • Updated • 23.7k • 112 -
unsloth/llama-3-8b-Instruct-bnb-4bit
Text Generation • 8B • Updated • 47k • 134 -
unsloth/gemma-2-9b-bnb-4bit
Text Generation • 10B • Updated • 10.4k • 31
-
unsloth/Llama-3.2-3B-Instruct-bnb-4bit
Text Generation • 3B • Updated • 57.1k • 36 -
unsloth/Llama-3.2-1B-Instruct-bnb-4bit
Text Generation • 1B • Updated • 40.8k • 23 -
unsloth/Llama-3.2-11B-Vision-Instruct-bnb-4bit
Image-Text-to-Text • 11B • Updated • 2.89k • 83 -
unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit
Text Generation • 8B • Updated • 68.2k • 102