I wasn't looking for this. I was digging through Claude Code skill repos at 1am, the way you do, and I almost scrolled past a small project called Guardsman. Two stars. Eighteen commits. No flashy benchmark chart at the top of the README.
That absence is exactly what made me stop scrolling.
The pattern I'm tired of
Every "AI coding skill" README follows the same script now: a bold claim in the title ("73% fewer tokens!", "10x faster shipping!"), a chart with no visible methodology, and a repo that's three weeks old. You can't reproduce the number. You can't even tell what it was measured against. You just have to believe it.
Guardsman does the opposite, and it says so out loud, in its own README:






