Comment on Surpassing Frontier Performance with FusionparentComments−andai2moI took a better look at their graphs. Their Opus+Opus fusion indeed matches Fable on this benchmark, but costs nearly twice as much as Fable!Meanwhile their "budget" fusion almost matches Fable, and costs half as much.At least on this benchmark. (Which looks a bit odd to me, e.g. DeepSeek ouranks GPT-5.5, ???)Would love to see more benchmarks testing this technique.
Comments
I took a better look at their graphs. Their Opus+Opus fusion indeed matches Fable on this benchmark, but costs nearly twice as much as Fable!
Meanwhile their "budget" fusion almost matches Fable, and costs half as much.
At least on this benchmark. (Which looks a bit odd to me, e.g. DeepSeek ouranks GPT-5.5, ???)
Would love to see more benchmarks testing this technique.