inclusionAI/Ling-3.0-flash

#2824
by piloponth - opened

Hello team,

the support for Ling-3.0 has been merged to llama.cpp (see https://github.com/ggml-org/llama.cpp/pull/26608)

Kindly asking for quanting the https://huggingface.co/inclusionAI/Ling-3.0-flash into .gguf of your famous quality.

Best,
Pilo.

CISC merged commit 3733366 into ggml-org:master yesterday

let's hope we are new enough =)
if not, touch me again please =)

It's queued! =)

You can check for progress at http://hf.tst.eu/status.html or regularly check the model
summary page at https://hf.tst.eu/model#Ling-3.0-flash-GGUF for quants to appear.

let's hope we are new enough =)

Please don't hope and queue it on rich1. rich1 worked a whole night on this, preventing meaningful quantisation of other models, for a model that clearly isn't supported yet - nico last merged our llama.cpp with upstream on the 15th, I think, so this cannot be included. Manually queue it for nico1 - nico1 has disk bandwidth to spare in case we need to wait and restart.

Also, feel free to bug nico. once he has merged it, our binaries usually get updated the next time i queue models.

sigh :)

Sign up or log in to comment