Skip to content

Comment on GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best?

Comments

A really frustrating partial presentation, given an apparent lack of testing with a spread of efforts for each model.

Given that there's no reason to believe that Fable's xhigh is comparable to GPT-sol's xhigh, or Opus xhigh, for that matter, it would be far more useful to see the effort level where these tasks no longer achieved their goals.

These benchmarks are done with incorrectly and missing a lot of baselines. The article seems very vibe coded too.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.