Storia: Weka's new storage platform caches 100% of a model's pre-calculated tokens, so it never has to redo the work — Warptech Lab News