MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents

📅 2026-08-06
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the challenge that long-horizon agents often fail to effectively reuse stored experiences due to representational mismatches between past memories and current task contexts. To overcome this, the authors propose MemPrism, a framework that decouples persistent memory from task-time working memory and dynamically constructs relational memory views conditioned on the current task. MemPrism employs a lightweight view policy, a deterministic composer, and a renderer to generate task-adapted ephemeral memory structures from event streams, enabling seamless transfer across diverse vision-language models without fine-tuning. Experiments demonstrate that MemPrism significantly improves performance on long-horizon embodied and web-based agent benchmarks, particularly excelling in extended-task scenarios while substantially reducing memory token consumption.
📝 Abstract
Long-horizon agents rely on memory to reuse experiences, yet existing memory systems often assume that evidence can be directly consumed through a fixed representation. This leads to representation mismatch, where relevant information is available but not organized for the current decision. To this end, we propose MemPrism, a task-conditioned relational memory framework that separates persistent experience storage from decision-time working memory. MemPrism records interactions as the event stream and dynamically constructs relational views according to the current task context. A lightweight view policy selects the relation structure, evidence range, outcome condition, and granularity, while a deterministic composer and render transform historical facts into a temporary optical working-memory view for a frozen task policy. Experiments on long-horizon embodied and web-agent benchmarks show that MemPrism consistently improves the task performance, especially as trajectories become longer, while reducing memory token consumption. Furthermore, the learned view policy transfers across different VLMs without additional adaptation, demonstrating the effectiveness of task-conditioned relational views as a general memory interface for agents.
Problem

Research questions and friction points this paper is trying to address.

long-horizon agents
memory systems
representation mismatch
relational memory
task-conditioned
Innovation

Methods, ideas, or system contributions that make the work stand out.

task-conditioned memory
relational views
long-horizon agents
working memory
memory efficiency
Z
Zhisheng Chen
Nanyang Technological University
B
Bingfan Zeng
South China University of Technology
B
Bangde Cao
Beijing University of Posts and Telecommunications
Z
Zhengwei Xie
University of Science and Technology of China
Y
Yuxuan Li
Nanyang Technological University
Jinhan Li
Jinhan Li
Undergraduate Student, New York University
Z
Zheng Lu
Peking University
X
Xiangchen Guan
Peking University
Z
Zikai Xiao
Zhejiang University
R
Rui Qian
Fudan University
Jingwei Song
Jingwei Song
University of Michigan
SLAM3D reconstructionSurgical visionGPU programming