Your personal memory, across sessions, agents, and devices.

Earlier-version benchmark

An Earlier Version of memU Achieved 92.09% on LoCoMo

This result is retained as a historical benchmark. It was measured on an earlier version of memU and should not be read as a benchmark of the current architecture.

LoCoMo benchmark result for MemU

About Long Conversation Memory

The LoCoMo (Long Conversation Memory) dataset is the industry‑standard benchmark for evaluating long‑term memory and reasoning in conversational AI systems. It is built from very long multi‑session dialogues with rich temporal, personal, and event‑driven context, enabling comprehensive evaluation across multiple reasoning categories such as single‑hop retrieval, multi‑hop inference, temporal understanding, and open‑domain question answering. Originally developed through a hybrid human–machine annotation process and adopted widely by the research community, LoCoMo provides a standardized framework to assess an AI’s ability to recall, reason, and persist information across extended interactions.

Explore the Current memU

See the current skill-driven memory framework, installation instructions, and source code on GitHub.