An Earlier Version of memU Achieved 92.09% on LoCoMo
This result is retained as a historical benchmark. It was measured on an earlier version of memU and should not be read as a benchmark of the current architecture.

About Long Conversation Memory
The LoCoMo (Long Conversation Memory) dataset is the industry‑standard benchmark for evaluating long‑term memory and reasoning in conversational AI systems. It is built from very long multi‑session dialogues with rich temporal, personal, and event‑driven context, enabling comprehensive evaluation across multiple reasoning categories such as single‑hop retrieval, multi‑hop inference, temporal understanding, and open‑domain question answering. Originally developed through a hybrid human–machine annotation process and adopted widely by the research community, LoCoMo provides a standardized framework to assess an AI’s ability to recall, reason, and persist information across extended interactions.
Explore the Current memU
See the current skill-driven memory framework, installation instructions, and source code on GitHub.