← Back to context

Comment by HarHarVeryFunny

9 hours ago

> They didn't publish the tools they used to train the model.

That's not totally true - for example Ziphu (Z.ai, developers of GLM) have published their Slime RL-training framework, developed together with Tsinghua University.

https://github.com/THUDM/slime

Also, if you read Moonshot's Kimi 3 report, it does give a lot of architectural detail and details on training.

https://github.com/MoonshotAI/Kimi-K3/blob/main/k3_tech_repo...

Some discussion on this by HuggingFace here:

https://www.youtube.com/watch?v=MW8-kqd2SD8

Yes, we all realize that "open weights" is more accurate than "open source", and anyways the source code would not be very interesting - it's the training data and methods that mostly define these models.