The agent proposes a plan. It sounds reasonable — they always do, because the agent is exactly as fluent describing a bad approach as a good one. You say go. Two hundred lines later the approach hits the wall it was always going to hit: the assumption nobody questioned, the edge it never accounted for. The plan was confident. Confidence was never the thing to check.
A plan you only ever hear arguments for is a plan you can't actually judge. The cheapest pressure-test is to make the same agent that proposed it try to take it down — before any code leans on it, while changing direction still costs a sentence instead of a rewrite:
✕ Without the prompt
✓ With the prompt
Paste it after any non-trivial plan, before the build. It costs one extra exchange and kills the most expensive kind of rework — the kind where the code was clean but built toward an approach that was wrong from the first line. You're not hunting for the cleverest plan. You're looking for the one that survives contact with the real problem.