🤗 Hugging Face   |   🤖 ModelScope    |   🐙 OpenRouter   

Ling-3.0-flash-GGUF

Ling-3.0-flash is our next-generation native hybrid reasoning model. Operating with 124B total and 5.1B active parameters (~12.4% and ~8.1% of our previous 1T-class flagship Ring-2.6-1T), Ling-3.0-flash matches or outperforms its predecessor across key benchmarks.

Find more details in the original model card: https://huggingface.co/inclusionAI/Ling-3.0-flash

Downloads last month
241
GGUF
Model size
127B params
Architecture
bailingmoe3
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including inclusionAI/Ling-3.0-flash-GGUF