Qwen3 4B Instruct 2507 β€” the best mobile model of 2026?

#30
by 3morixd - opened

We've been benchmarking Qwen3-4B-Instruct-2507 on our phone farm. The results are impressive:

  • 12.4 t/s on Snapdragon 865 (Q4_K_M)
  • 2.4GB model size
  • Excellent reasoning quality
  • Strong code generation
  • Good Arabic support

This might be the best "Goldilocks" model β€” big enough for real tasks, small enough for mobile. We're considering it as our default deployment model.

The July 2025 update improved instruction following significantly. Qwen keeps getting better.

β€” Dispatch AI (FZE), Sharjah UAE

Sign up or log in to comment