A benchmark result that changes what we thought was possible for local persistent agent vector memory

We ran VEKTOR Slipstream against LongMemEval this week and got a result we were very impressed with.

79.0%. That is 12 points above full-context GPT-4, 17 above Mem0, 24 above ReadAgent, and 30 above MemGPT.

To understand why that number matters, you need to understand what LongMemEval is actually testing, why it is hard, and what it took to get there.

What LongMemEval Is and Why It Is the Hardest Memory Benchmark