Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

TiMem: Temporal-Hierarchical Memory Consolidation for Long-Horizon Conversational Agents

About

Long-horizon conversational agents have to manage ever-growing interaction histories that quickly exceed the finite context windows of large language models (LLMs). Existing memory frameworks provide limited support for temporally structured information across hierarchical levels, often leading to fragmented memories and unstable long-horizon personalization. We present TiMem, a temporal--hierarchical memory framework that organizes conversations through a Temporal Memory Tree (TMT), enabling systematic memory consolidation from raw conversational observations to progressively abstracted persona representations. TiMem is characterized by three core properties: (1) temporal--hierarchical organization through TMT; (2) semantic-guided consolidation that enables memory integration across hierarchical levels without fine-tuning; and (3) complexity-aware memory recall that balances precision and efficiency across queries of varying complexity. Under a consistent evaluation setup, TiMem achieves state-of-the-art accuracy on both benchmarks, reaching 75.30% on LoCoMo and 76.88% on LongMemEval-S. It outperforms all evaluated baselines while reducing the recalled memory length by 52.20% on LoCoMo. Manifold analysis indicates clear persona separation on LoCoMo and reduced dispersion on LongMemEval-S. Overall, TiMem treats temporal continuity as a first-class organizing principle for long-horizon memory in conversational agents. The code is available at https://github.com/TiMEM-AI/timem.

Kai Li, Xuanqing Yu, Ziyi Ni, Yi Zeng, Yao Xu, Zheqing Zhang, Xin Li, Jitao Sang, Xiaogang Duan, Xuelei Wang, Chengbao Liu, Jie Tan• 2026

Related benchmarks

TaskDatasetResultRank
Long-context Question AnsweringLocomo--
174
Long-context Question AnsweringLocomo--
45
Long-term memory evaluationLongMemEval S (test)
KU (Knowledge Update)87.69
30
Long-term MemoryLongMemEval
Score74.9
30
Memory-augmented Question AnsweringLocomo
Accuracy (Temporal)77.63
25
Person understanding and persistent memoryLoCoMo-Plus
Score24.2
21
Person understanding and persistent memoryRealMem
Score46.1
21
Person understanding and persistent memoryKnowMe
Score52.1
18
Person understanding and persistent memoryRealPref
Score87.9
18
Person understanding and persistent memoryCUPID
Score60.6
18
Showing 10 of 16 rows

Other info

Follow for update