← Back to context

Comment by dofm

7 days ago

FWIW my tests on my little puzzles suggest that it is not better than the Gemma 4 12B on SQL. It really does seem to get quite tangled up on stuff.

PHP/Wordpress code seems OK (better than the Gemma) but it gets stuck in reasoning loops.

Mind you, I am something of a cynic about the underlying 27B dense Qwen; I think the 35B MoE model is often better and it is just so, so much faster.

Qwen3.6-27B is the best model in that range that I’ve used for agentic coding by far. I think it’s kinda mid at everything else.

  • Surely not that good at vision. TBH none of these 14-27b models come close to even the cheapest Gemma 2.5 flash.

    If these buddies are similarly bad on text, then they definitely don’t get anywhere close to big boys, no matter what the synthetic stats claim upon release.