Storia: Stop Wasting LLM Budgets: High-Performance Semantic Caching with Spring AI and pgvector — Warptech Lab News