Comment on Kimi K1.5: Scaling Reinforcement Learning with LLMsparentComments−m00x1ywdym? They did their own pretraining.
Comments
wdym? They did their own pretraining.