The second look — reflection and self-correction.
The best move in the loop is often noticing a result can't be trusted — and not building on it.
Not every step goes cleanly. A tool returns nothing useful, a reading is ambiguous, an early guess turns out wrong. What separates a robust agent from a brittle one is what it does next: charge ahead on a shaky result, or take a second look?
Reflection is the agent examining its own last step — is this reliable enough to build on? — and, when it is not, changing course: retry, reach for a different tool, or revise an earlier call. This is where seeing the outcome pays off: the agent can catch a weak step and try again, where a one-shot answer is simply stuck with it.
Same shaky reading, two agents. Flip between them and watch where the ruling ends up.
Doubt the weak step, don't build on it.
Both agents got the same unreliable reading. One treated it as fact and ruled on sand; the other doubted it, refused to conclude, and reached for a better instrument. That extra reflect step — questioning its own last result — is the whole difference between the wrong call and the right one.
Self-correction is not second-guessing everything; it is spotting when a result is weak, a tool failed, or an assumption broke, and adjusting before the mistake compounds. An agent that never looks twice fails the moment a single step misleads it.