🤗 Hugging Face   |   🤖 ModelScope    |   🐙 OpenRouter   

Ling-3.0-tiny-GGUF

Ling-3.0-tiny is a lightweight hybrid reasoning MoE model with 7.9B total parameters and only 1.3B activated parameters per token. It is designed to deliver strong reasoning and agentic capabilities at low inference cost, making advanced model capabilities more accessible for local and resource-constrained deployment.

Find more details in the original model card: https://huggingface.co/inclusionAI/Ling-3.0-tiny

Downloads last month
5
GGUF
Model size
8B params
Architecture
bailingmoe3
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including inclusionAI/Ling-3.0-tiny-GGUF