A maintainer recently invited me to test AgentInspect, a local-first toolkit for debugging and testing TypeScript AI-agent trajectories.
I was interested, but a proper review could easily turn into hours of setup, testing, documentation, and reproduction work. Instead of choosing between a superficial comment and a large manual audit, I tried a third option: an AI-assisted black-box test with clear boundaries.
The result was a reproducible bug report that the maintainer confirmed and fixed.
The testing rule that mattered most
The AI agent was instructed to behave like a new external user.






