Preprint
Jul 2026
Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation
LaMem-VLA is introduced, a latent-memory-native framework that reconstructs historical experience into latent memory tokens and directly interweaves them with VLA reasoning, and enables memory to directly participate in VLA reasoning and guide action generation under a bounded context.
Hongyu Qu, Jianzhe Gao, Xiaobin Hu et al.
· 1 citation