Comment on AI researchers debate how close we are to recursive self-improvementparentComments−aix11hAre you saying the models are already autonomously constructing next-gen evals for themselves? (Which is what the GP is asking.)−nl1hSure?Doesn't everyone get their agents to construct evals it can't pass? There's nothing magical about this.−aix135mWould love to learn more about some techniques that "everybody" uses to do this well. So far, everything I've seen that meaningfully advances the frontier has been high-touch (involving human experts in one way or another).
Comments
Are you saying the models are already autonomously constructing next-gen evals for themselves? (Which is what the GP is asking.)
Sure?
Doesn't everyone get their agents to construct evals it can't pass? There's nothing magical about this.
Would love to learn more about some techniques that "everybody" uses to do this well. So far, everything I've seen that meaningfully advances the frontier has been high-touch (involving human experts in one way or another).