Memory recall accuracy
90%
answers correct on 30 memory tests
We ran 30 real-world memory scenarios, remembering things said in one conversation, across multiple sessions, and at specific points in time. An independent judge scored each answer. Single-session recall scored 100%; multi-session and time-based recall averaged 83%.
Measured with LongMemEval, an industry-standard test with a live judge. 90% is near current state-of-the-art.
Smarter retrieval
+75%
more accurate than dumping all context in
When Cognition retrieves structured memory instead of pasting everything into the prompt, answers are 75% more accurate, and the agent processes 39% fewer tokens to get there. Less noise, sharper signal.
Measured with LoCoMo. Validates that structured memory beats raw context stuffing.
Automated code repair
0%
honest early pilot result
We tested memory-assisted automated code repair on 2 real tasks. The memory layer retrieved correctly, but a downstream technical step hit a blocker before scoring could run. We're showing the real number, not a polished one.
Measured with SWE-Bench-CL. Memory infrastructure is live. We'll update this as the pipeline matures.