StateMem: Single-State Residual Memory with Adaptive Inference for Vision-Language-Action Policies
Memory-dependent robotic manipulation often requires later actions to use information from earlier interactions. Existing vision-language-action (VLA) policies primarily rely on current observations, limiting historical information retention. Memory-augmented VLAs, such as MemoryVLA, address this limitation with extern...