Skip to content

Comment on Introducing the Open Chain of Thought Leaderboard

Comments

Self-discovery > Self-consistency > Medprompt > Chain of thought

I think the leaderboard the way it is devised is a bit silly, it rewards failure of the base model and success of the prompt atop it, but that is not how we want to be using the style of prompting. We need to see it how the gorrila code nudging metric does it, both base model score and the increase from the prompt style matter.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.