Qwen3.8 35B A3B model, please, please, please.

#43
by Duonglv - opened

Thank Qwen so much and so much for your contributions.

I hope that you will release the MOE model, maybe Qwen3.8 35B A3B.
I think there are many many people need it like me.

The 27B dense model focuses on the local tasks.
While Qwen3.8 35B A3B is much better for speed to serve many concurrent requests.

The pair 27B Dense and 35B Moe model is perfect.

Yeah, Qwen 3.8 35B A3B weight is really essential for those of us who don't have much VRAM

Also a 7B to 14B model please! The previous Qwen models have pretty good size ranges, would be very nice to continue that trend!

+1 But also ensure Qwen 3.8 35B A3B is instruction tuned 3.6 sure loved going off rails

I tried the 27B, it can fit the 5090 in theory but at least on my rig on specific unsloth studio which is the only inference app that works properly with multi-agent harnesses - at least on my system, it driver TDR's my 5090 because of those dense layers, I really need the MOE 35B version.

We need a 35b ish mor model which can keep the speed and performance balance. please

i vote for it

i vote for it X2
regards

I like sparse models

Sign up or log in to comment