back

by enraged_camel·1mo ago·view on hn ↗
The charts are also extremely difficult to parse. They seem auto-generated. Dataset coloring is atrocious.

Regarding your main point, yes, I agree. My impression (as someone who uses both Codex and Claude Code daily) is that OpenAI does a fair amount of benchmaxxing.