I gave two big coding tasks to both Claude and Codex.

Claude finished in about one hour. Codex took about eight.

That sounds like a clean win for Claude until you look at what came back. The Claude output was fast, confident, and useless. Bad assumptions, broken code, missing integration points, half-followed rules, and a shape that would make...