← Back to context

Comment by intothemild

5 hours ago

So they trained a model on open weights, and then aren't releasing the weights... am I reading this right?

Technically kimi k-3 weights license is not open weight (it has a lot of restrictions). I would classify it as ‘weight open’ similar to the bsl and fsl ’source open’ licenses.

There is little to no point reading the article as well. It's stripped of all alpha.

> task and environment feedback

> on-policy planning and learning

> feedback connects decisions to their consequences

These are deliberately the least informative phrases you could possibly use to describe what you have done, while still being in the realm of words that go over a generic investor who has no idea whats going on and may be dazzled by sciencey sounding language.

Cursor compose 2.5 article where they used and described on policy self distilation was actual alpha.

Aren't Cursor Composer models like this too? At some point all the extra RL you do can be considered as proprietary information added.

Not suggesting this is right or wrong, but is sort of the nature of the technology.

It happens with open source software all the time, why would we expect any different with open source weights.