From GitHub
- GitHub stars
- 849
- Last code update
- 2026-08-07
A fixed evaluation harness and leaderboard for comparing LLM agent memory systems.
Why it made the shelf
It is an open-source evaluation harness and tool for testing and benchmarking LLM agent memory systems.
Seen on
Official signals
Sourced facts from the places where this tool ships.
From GitHub
Public trail
1 reference from 1 publisher.