Skip to content

Comment on Show HN: LlamaGym – fine-tune LLM agents with online reinforcement learningparent

Comments

I’m not really sure what your point is. Is it not remarkable that valuable things can be done in 150 lines?

I agree with you. / Above, I wouldn't assume a single nor clearly intended "point". Reading it I got an impression more of concern, even fear. I'm guessing one underlying driver may be a concern that AI is creeping into more and more programming. Which is true.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.