Developer builds AI memory system Muninn, misses leaderboard deadline but self-benchmarks results
A developer and their AI partner built a hybrid memory retrieval system called Muninn overnight to enter the Agent Memory Leaderboard, which pits systems against competitors from Tencent, Mem0, Cognee, and MemOS. The team missed the submission window and will try again when the next cycle opens in September. Running the benchmark's public pipeline independently on the LoCoMo dataset, Muninn scored an estimated 72.9% in its best configuration, though the developer cautions this is an internal estimate rather than an official result. The same core system, entered as Perpetual Recall on the separate LongMemEval-V2 benchmark, achieved a confirmed submission score of 56.98% accuracy with a query latency of roughly 2.3 seconds. While mid-pack on accuracy, every system that outperformed it required between 27 and 180 seconds per query, compared to under three seconds for Perpetual Recall.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in