max_position_embeddings/context_length incorrect in GGUF headers?

#1
by dinerburger - opened

The source config.json indicates the max_position_embeddings to be 256K, but this appears to be clamped at 128K. Was there a technical reason for the change? Or is this something to do with the converter script? (PS: thanks again for your hard work bringing this super useful model into the GGML fold!)

Oh gosh sorry I see now they updated the value two weeks ago my apologies.

dinerburger changed discussion status to closed

You were correct. MXFP4_MOE, Q1_0, IQ1_S, IQ1_M, IQ2_M, IQ3_XXS, UD-Q2_K_XL, Q3_K_M, Q4_K_M, Q5_K_M, Q5_K_S, Q6_K, Q8_0, and UD-Q6_K_XL have all been corrected and re-uploaded with 262k. BF16, Q4_K_S, and UD-Q4_K_XL and in process now, and will be soon to follow with the next hour or so (check file timestamps, if it is today it is updated).

Yea I just ended up modifying the KV meta values but reuploading'll help other folks looking to get going with the model.

dinerburger changed discussion status to open
dinerburger changed discussion status to closed

Sign up or log in to comment