Skip to content

Comment on Every Model Cheatsparent

Comments

Except these AIs are both scary and smart. I think the characterization is warranted.

The whole point of calling it “cheating” is to warn us that if we tell an AI to do something, we need to take into account that it may well do something we don't expect; something that we as humans dismiss as “cheating”.

Someone in this thread gave a good example: if you ask an AI to get a reservation at a restaurant, and make sure the reservation is in place before exiting the agent loop, you don't expect the AI to hack the restaurant and erase someone else's reservation to make space for yours. But an AI will absolutely do that.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.