Comment by cobanov

5 hours ago

Developer here. You're right, Laya is a lot weaker than Jev, especially on harder queries. It's a small model, so it's fast, but that's the trade-off. The open models that get close to Jev are much bigger, and running those is what I'm working on next.

It doesn’t to be a ton bigger, 16k and reliable 8k would be a godsend. (I run at 2k)

What are the models? I am super curious in these as well

  • Probably Kev and/or the decider models. Kev is trained on one of the 4B qwen models, similar for decider but it ranges from 0.8B through to the 35B-A3B model so far I believe.