|
|
|
|
|
by enraged_camel
21 days ago
|
|
The charts are also extremely difficult to parse. They seem auto-generated. Dataset coloring is atrocious. Regarding your main point, yes, I agree. My impression (as someone who uses both Codex and Claude Code daily) is that OpenAI does a fair amount of benchmaxxing. |
|