← Back to context

Comment by trvz

7 days ago

I still don’t see the point of this. In my testing, it’s worse than Qwen 3.5 4B and even 0.8B.

When new models are released (I realized this is qwen 3.6 but the quant is novel) - it takes a few days for the kinks to get worked out - you’ll likely have better luck if you give it a few days and try again.

It is still worth experimenting. Although we do need independent evaluation of these models instead of labs posting such biased and skewed results.