Much better with the original chat template

#3
by Neiko2002 - opened

These are the benchlocal benchmark results with the chat template of this repo:
image

And I get these numbers when using the original chat template:
image

TeichAI org

Interesting! What are the benchmarks measuring?

Those numbers are https://benchlocal.com/ scores. Here is another version with "empero-ai/Qwythos-9B-Claude-Mythos-5-1M" chat template. A template which sometimes improves cli and dataextract, but slightly degrades other tests. FYI base Qwen 9B reaches an average score of 76.9.

image

Sign up or log in to comment