๐Ÿค— Hugging Face   |   ๐Ÿค– ModelScope    |   ๐Ÿ™ OpenRouter   

Ling-3.0-flash-GGUF

Ling-3.0-flash is our next-generation native hybrid reasoning model. Operating with 124B total and 5.1B active parameters (~12.4% and ~8.1% of our previous 1T-class flagship Ring-2.6-1T), Ling-3.0-flash matches or outperforms its predecessor across key benchmarks.

Find more details in the original model card: https://huggingface.co/inclusionAI/Ling-3.0-flash

Downloads last month
8,816
GGUF
Model size
127B params
Architecture
bailingmoe3
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Space using inclusionAI/Ling-3.0-flash-GGUF 1

Collection including inclusionAI/Ling-3.0-flash-GGUF