AI memory has a measurement problem.
"How much context?" tells us less and less.
What matters is what an agent remembers, forgets, and uses on the next task.
@MemoraX_AI taking #1 in the first Agent Memory Leaderboard Commercial Products, Text Memory track is a real milestone.
It traces failures across memory writing, organization, retrieval, reranking, fusion, and memory use. Those task outcomes can then feed into strategy updates and regression evaluation.
In other words, memory itself becomes something you can observe, modify, and retest.
That feels like a healthier direction for the category.