Hacker News new | ask | show | jobs
by enraged_camel 21 days ago
The charts are also extremely difficult to parse. They seem auto-generated. Dataset coloring is atrocious.

Regarding your main point, yes, I agree. My impression (as someone who uses both Codex and Claude Code daily) is that OpenAI does a fair amount of benchmaxxing.