Comment by yassa9

2 days ago

Can anyone who has that specific personal test he tries on different models , and tries this model , to tell us here if possible , how good or bad is this new model ? compared to others ?

I only trust those users genuine personal tests

There is a down to earth guy on YT that performs a series of tests against LLMs running on non-god-tier commodity hardware. He will likely be testing this soon enough.

https://www.youtube.com/@lukesdevlab

I don't know if that is what you are looking for or not and as always your experiences may be different.

  • So much potential for that channel. He's got a nice range of tests and a no nonsense presentation style.

    However, watching tests of heavily quantized models that weren't designed for it (non-QAT) is frustrating. There's no way to tell if the actual model fails because it's dumb or if the lobotomy made it that way.

    • I noticed he does pay attention to feedback on his videos and I think some people have pointed that out.

  • thaaanks man, this channel seems really informative, although < 10K subs only !

    • It's a relatively new channel - but yeah - I feel the guy puts a lot of effort into what he does and deserves more subs.

yeah this model is chefs kiss..

running an untouched, vanilla 4-bit version (Q4_0) I baked myself today (benched it against Q4_K_M (16gb) and IQ3_M (12gb), Q4_0 (15gb) is king)...

this model--

1: over 60% faster than qwen 3.6 version of the same dense 27b model, same engine setup (don't ask how, im not sure either)

2: has better reasoning quality, less "loopy" with its thinking patterns.. most certainly the smartest model on my roster currently

3: has the longest task horizon ive ever experienced (locally or otherwise)...I sent it a bunch of compressed ideas for an app, it sent me back the largest python app ive ever seen in one single ai response pass (80kb text file)

thanks qwen!! hoping to see the full model range get released...