Add in-depth cross-paper review of VLA memory architectures
Covers 18 papers spanning π-series MEM, ICLR 2026 (MemoryVLA, HAMLET,
Ctrl-World), CoRL 2025 (Long-VLA), and recent 2024-2026 preprints
(MemER, SAM2Act+, ContextVLA, CronusVLA, ReMem-VLA, EchoVLA,
ExpReS-VLA, Keyframe-Chaining VLA, MAP-VLA, TraceVLA, LoHoVLA,
EvoVLA, RoboMemory).
Page structure:
- Why memory matters; HAMLET's 29.2 -> 76.4% headline delta on GR00T N1.5
- Six-category taxonomy (compressed video tokens, cross-attention
memory banks, symbolic/cognitive memory, retrieval-based episodic
stores, raw-frame history, planning-token hierarchies)
- Per-paper deep-dives with arXiv links
- Cross-axis comparison table (cheapest plug-in, most interpretable,
longest horizon, best for spatial / on-device / hierarchical tasks)
- Trends: from stateless (2023) to hybrid multi-scale (2026);
plug-and-play on frozen backbones; retrieval eating robot memory
- Open questions: selective write, hour-scale horizon, episodic vs
parametric, interpretability probes, standard benchmark gap
- Practical decision guide for shipping VLAs
Cross-links added in HAMLET, MemoryVLA summary pages; sidebar and
Home updated to surface the new review.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>