I think this is the key takeaway from here.
Meanwhile their "budget" fusion almost matches Fable, and costs half as much.
At least on this benchmark. (Which looks a bit odd to me, e.g. DeepSeek ouranks GPT-5.5, ???)
Would love to see more benchmarks testing this technique.
Meanwhile their "budget" fusion almost matches Fable, and costs half as much.
At least on this benchmark. (Which looks a bit odd to me, e.g. DeepSeek ouranks GPT-5.5, ???)
Would love to see more benchmarks testing this technique.