Got this great quote from Garry Kasparov in Wired's article on multi-agent RL[1]:
“Creativity has a human quality. It accepts the notion of failure."
As faithful min-maxers, LLMs are always going to have an overconfident Prisoner's Dilemma blind spot in their algorithms. Unlike their cinematic brethren, they're progammatically unable to conclude with "the only winning move is not to play."
This seems like the next major hill to conquer to make them useful.
Comments
Got this great quote from Garry Kasparov in Wired's article on multi-agent RL[1]:
As faithful min-maxers, LLMs are always going to have an overconfident Prisoner's Dilemma blind spot in their algorithms. Unlike their cinematic brethren, they're progammatically unable to conclude with "the only winning move is not to play."
This seems like the next major hill to conquer to make them useful.
[1] https://www.wired.com/story/google-artificial-intelligence-c... - kind of a meh article otherwise