Skip to content

Comment on Task-free intelligence testing of LLMsparent

Comments

Doesn't that presume that one model dominates the other?

It presumes some models are better than others (and we do find that providing data with a wide mix of model strengths improves convergence) but it does not need to be one model, and it does not even need to be transitive.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.