So they trained a model on open weights, and then aren't releasing the weights... am I reading this right?
It happens. Most open licenses aren't GPL style copyleft.
There is little to no point reading the article as well. It's stripped of all alpha.
> task and environment feedback
> on-policy planning and learning
> feedback connects decisions to their consequences
These are deliberately the least informative phrases you could possibly use to describe what you have done, while still being in the realm of words that go over a generic investor who has no idea whats going on and may be dazzled by sciencey sounding language.
Cursor compose 2.5 article where they used and described on policy self distilation was actual alpha.
Aren't Cursor Composer models like this too? At some point all the extra RL you do can be considered as proprietary information added.
Not suggesting this is right or wrong, but is sort of the nature of the technology.
It happens with open source software all the time, why would we expect any different with open source weights.
Which is fine, that’s legal according to the license
Technically kimi k-3 weights license is not open weight (it has a lot of restrictions). I would classify it as ‘weight open’ similar to the bsl and fsl ’source open’ licenses.