why is this model faster than Qwen3.5 9B

#16
by dpe1 - opened

I donwloaded Q6 and it is 70 tokens per second, when Qwen3.5 unsloth Q4 something is 20 tokens per second

Qwen 3.5 is general purpose (does everything but doesn't have a domain where it is better)
while omnicoder is coding purpose

Sign up or log in to comment