MTP?

#1
by Trilogix1 - opened

Any plan to mtp?

@Trilogix1 I think you missed something :)

image

@ilintar

Good catch thanks (appreciated) , this happens when you have 100s of version. I used an old build to quantize :)

Hugston-Macaron-V1-Tall-Qx_K_M.gguf .... load_model: failed to create MTP context....

I wouldn't recommend running this on llama.cpp to be honest. Without the LoRA routing this is just a plain old Qwen3.6 35B-A3B. There's another model currently being ported with a similar architecture (Granite-Switch), so once that gets merged I think we can do this one more easily.

Sign up or log in to comment