What do you actually check when reviewing code your agent wrote?
A developer asks what people genuinely verify when reviewing agent-written pull requests, since everything looks finished: green tests, confident descriptions, easy merges. The thread converges on reviewing the agent's silent decisions (defaults, skipped edge cases, self-granted permissions) and diffing against the original request rather than the PR description.