A maintainer recently invited me to test AgentInspect, a local-first toolkit for debugging and testing TypeScript AI-agent trajectories.

I was interested, but a proper review could easily turn into hours of setup, testing, documentation, and reproduction work. Instead of choosing between a superficial comment and a large manual audit, I tried a third option: an AI-assisted black-box test with clear boundaries.

The result was a reproducible bug report that the maintainer confirmed and fixed.

The testing rule that mattered most

The AI agent was instructed to behave like a new external user.