Comment on GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best?parentComments−pohl1moSeemed strange they would bench Terra and Luna but not Opus, Sonnet, and Haiku−StefanKarpinski1moThey were all benchmarked, but not included in the post to try to keep the amount of data from being excessive. (Results are also not surprising — Opus good, Sonnet struggles, Haiku fails.)
Comments
Seemed strange they would bench Terra and Luna but not Opus, Sonnet, and Haiku
They were all benchmarked, but not included in the post to try to keep the amount of data from being excessive. (Results are also not surprising — Opus good, Sonnet struggles, Haiku fails.)