Skip to content

Comment on Kimi K1.5: Scaling Reinforcement Learning with LLMs

Comments

The set of math/logic problems behind AIME 2024 appears to be... https://artofproblemsolving.com/wiki/index.php/2024_AIME_I_P...

Impressive stuff! But unclear to me if it's literally just these 15 or if there's a large problem set...

The full dataset is here - https://huggingface.co/datasets/AI-MO/aimo-validation-aime you can use the eval script I have in optillm to benchmark on it - https://github.com/codelion/optillm/blob/main/scripts/eval_a...

doesn’t seem too hard to me, shame i was never exposed to this stuff in highschool

e: oh i see, they get progressively harder

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.