Good question, anybody know? I've seen harnesses optimize for token efficiency (maki), simultaneous agent operations (jcode), extensibility (pi), etc... But I have no idea if each focus area is superior and if the rest of the operations are comparable.
It bothers me that the Claude Code leak (from what I read, didn't look) a few months ago showed a lot of careful consideration for context handling. I assume this means that each harness performance/correctness would differ drastically just from this alone.
I’m with jpeeler here and I don’t think there’s a single optimum harness. Pi is really nice in that it allows you to hone in on what suits your needs, though. Some people won’t want to manage their harness though, so for them it’s probably not ideal.
All I can say with certainty is that Claude Code isn’t a panacea, and it’s quite buggy and clunky. I think this is evident just from regular usage, but becomes even clearer when you use harnesses that get in the way less and demonstrate better implementations across the board. Sometimes I think Claude Code feels like software that happened while other harnesses feel like they were engineered intentionally.
It bothers me that the Claude Code leak (from what I read, didn't look) a few months ago showed a lot of careful consideration for context handling. I assume this means that each harness performance/correctness would differ drastically just from this alone.