I stopped letting Claude Code review its own work
I’ve been testing a simple workflow: Claude Code writes or edits the code. Then, if the task is risky or messy, I hand it to Codex with one job: find bugs, bad assumptions, edge cases, and anything I should not ship. Across 53 review runs, Codex found meaningful issues 88% of the time. Total so far: 119 issues across 7 projects. The interesting part is not “Codex is better than Claude.” It’s that different models seem to miss different things. Claude is good at moving the work forward. Codex has been useful
评论
?
参与讨论