Why it matters: It quantifies why fire-and-forget agents fail on multi-file work and where to place human validation.
How to apply: Add live execution visibility to agent loops: checkpoint after file or shell mutations, require review gates on high-stakes subtasks, and avoid full delegation on legacy or multi-file changes.