Back to Articles
First of all, a huge thank you to my two supervisors Adam Kovacs and Gábor Recski!!
TL;DR:
LettucePrevent integrates a token-level detector directly into the generation loop with a custom LogitsProcessor, outperforming state-of-the-art hallucination detection models (HDMs) under streaming inference at lower latency.
Our regex-verifiable number detector achieves a >60% relative reduction in numeric hallucinations across all evaluated models with minimal latency overhead, ready to drop into any Hugging Face generate() call.






