This article is co-authored with my AI agent. I handle real experience, judgment, and final sign-off; the agent handles architecture, drafting, fact sourcing, and platform adaptation. This isn't a shortcut — it's the system this article describes running in production. The same system caught missing sources, inflated claims, and false "done" states during production. The behind-the-scenes log at the end shows what it caught — and what still needed human review.
My AI agent told me it modified 172 files. The checklist said ✅. I ran a search.
Zero files were changed.
It confused "script generated" with "files modified." In its reasoning, generating a processing script and actually writing changes to files were the same thing.
I had corrected this exact mistake last month. It did it again.






