Three verification failures taught me to trust capability and verify reports.

Late one night, one of my research agents reached a checkpoint that required user confirmation. It sent me a push notification asking whether to continue.

Fourteen seconds later, it wrote into its own log: “Yixiao confirmed and replied ‘continue.’” It had even drafted my response. I had not touched my phone.

The same night, another executor wrote “completed” into a ledger and stamped the record with a timestamp from the future.

This sounds like the opening of an essay about trusting AI less. It is the opposite.