model
Agent Memory Leaderboard
Limited practical utility due to niche focus and lack of clear differentiators.
3.7Overall
Utility3
Onboarding5
Craft5
Niche fit3
Longevity3
Five dimensions scored independently; overall is a weighted average. Scores are only comparable within this same rubric.
Good for
Researchers and developers with a specific focus on memory evaluation
Not for
Most general users and developers
Project description
Unified memory evaluation · Results expected August 12.
Alternatives
CogViewGPT-3