|
|
|
|
|
by hmokiguess
1 day ago
|
|
I said optimize your harness, not your model. My comment was with regards to Claude Code. There are lots of ways to achieve the same task with less turns with different engineered harnesses around more efficient models. Look at features they release, such as Dynamic Workflows, that spin up 100+ sub agents then reconcile the result. Does that make sense now? |
|
Now with that said, opensi's ultra mode is absolute trash - they should just have stolen Claude code's implementation - and there are clearly modes like max reasoning where they'll burn double the tokens to get another 10th of a percent performance to win benchmarks, which you should basically never be using. They're not making those versions at the expense of more efficient reasoning levels though and I don't mind that they exist (as long as they don't get made the default mode - openai - fix your ultra mode already).