Comment on Does RL Incentivize Reasoning in LLMs Beyond the Base Model?parentComments−seertaak1yThe authors of the paper address this argument in the QA section.
Comments
The authors of the paper address this argument in the QA section.