Request: GGUF quantization for inclusionAI/Ling-3.0-tiny

#1
by stornic56 - opened

Hi Bartowski!

First of all, thank you for all the fantastic GGUFs you provide to the community.

I wasn't sure where to send this, but I wanted to inquire about the possibility of requesting a GGUF conversion for the recently released Ling-3.0-tiny model (7.9B, 1.3B active, using the new bailingmoe3 architecture with KDA/MLA and Q-LoRA layers).

Since you've already done excellent work with previous models in the Ling family, I'd like to know if this request is appropriate.

Thanks in advance for all your work! uwu

Absolutely, and looking forward to it as soon as support is added!

https://github.com/ggml-org/llama.cpp/pull/26608

looks like the PR is getting to the end-game, so hopefully it won't be too long now :)

Thanks a million for everything you do for the community, Bartowski
I'll be keeping an eye out for it. Cheers :D

Absolutely, and looking forward to it as soon as support is added!

https://github.com/ggml-org/llama.cpp/pull/26608

looks like the PR is getting to the end-game, so hopefully it won't be too long now :)

Merged πŸ₯³πŸ₯³πŸ₯³

Sign up or log in to comment