Ask a guardrails library how often it is wrong and you will usually get silence, or a benchmark of how fast it is.
jamjet-guardrails ships nine deterministic checks for LLM input and output with no runtime dependencies. The nine checks are not the interesting part.
Every one of them publishes a precision and recall figure against a corpus committed to the repository, gated in CI. Change the detector and the numbers change, and the build notices.
The part I care about more: the cases it gets wrong are named, in the README, by case id.
The problem with a float







