Are you actively using this?

#1
by qenme - opened

I'm curious how you liked the model? Ornith seemed to be a bust, and Q3.5 397b is too old and never got updated. Did this quant work well for you?

I'm not actively using because qwen 3.5 122b and deepseek v4 flash both the best intelligence/efficiency ratio for my hardware, but I like this model a lot.

I have a few private benchmarks that only this model can solve consistently across all local models I was able to run locally. There's something special about its caveman reasoning that makes it stand out, but one downside is that it thinks a lot and every once in a while gets stuck in a reasoning loop.

IDK if the infinite reasoning loop is due to quantization, but I tried nex n2 mini without quantization (bf16) and happened there too. I wish they'd build a version of this model on top of qwen 3.5 122B.

I'm not actively using because qwen 3.5 122b and deepseek v4 flash both the best intelligence/efficiency ratio for my hardware, but I like this model a lot.

I have a few private benchmarks that only this model can solve consistently across all local models I was able to run locally. There's something special about its caveman reasoning that makes it stand out, but one downside is that it thinks a lot and every once in a while gets stuck in a reasoning loop.

IDK if the infinite reasoning loop is due to quantization, but I tried nex n2 mini without quantization (bf16) and happened there too. I wish they'd build a version of this model on top of qwen 3.5 122B.

I see. Thanks for the perspective and quant!

Sign up or log in to comment