Back to Research papers
Research paper index

OnEvoMemory: Evolving Memory through Online Robot Rollouts for Pretrained Robot Policies

Zhongxi Chen, Shenqi Zong

arXiv:2608.08749Published August 9, 20260 citations
  • cs.RO
  • manipulation
  • trajectory
  • action
  • policy
  • robot

Abstract

Long-horizon robot manipulation requires policies to track completed subtasks and critical interaction events. However, existing memory mechanisms heavily rely on external models or predefined update rules. To address this, we propose OnEvoMemory, a value-guided memory module for pretrained robot policies. It maintains recent context, high-value experiences, and salient transitions, while learning which experiences should be retained from trajectory outcomes. Offline demonstrations initialize the memory prior, whereas successful and unsuccessful online rollouts refine memory selection, helping the policy recognize task-stage transitions and avoid repeating completed subtasks. Experiments on long-horizon manipulation benchmarks show that OnEvoMemory improves the performance of the base VLA policy through both offline initialization and online memory evolution.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.