Skip to content

History / Review VLA Memory

Revisions

  • IROS 2026: add full-paper pages for RoboSSM and TempoFit; link from survey + reviews - IROS-2026-RoboSSM (KAIST + UT Austin/Peter Stone): scalable in-context imitation via state-space models (Longhorn SSM, linear-time, length extrapolation, LIBERO, open-source). The in-context/memory efficiency answer. - IROS-2026-TempoFit (XJTLU): training-free temporal retrofit that reuses a frozen VLA's prefix-attention K/V as content-addressable memory (layer-wise FIFO K/V + Frame-Gap Temporal Bias); LIBERO-Long +4.0%, near-real-time. - Linked both from the IROS survey (§3.3, §5.2), Review-In-Context-Imitation, and Review-VLA-Memory (structured-vs-parametric row). Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Sep 5, 2026
  • IROS 2026: add 5 full-paper analyses + reflect into existing in-depth reviews Full-paper pages (abstract-verified from official program): - IROS-2026-AtomVLA: subtask-aware VLA + latent-WM scoring for offline GRPO (WAM lens) - IROS-2026-3D-FlowMatch-Actor: CMU/NVIDIA unified single/dual-arm 3D policy, +41.4% PerAct2, ~30x faster (bimanual SOTA) - IROS-2026-EquiBim: symmetry-equivariant bimanual policy - IROS-2026-IMLE-VLA: single-step cIMLE action head, 55Hz, LIBERO 98.0% (efficiency) - IROS-2026-ICLR-Visual-Reasoning: in-context imitation with image-space reasoning traces Reflected IROS 2026 into existing reviews: - Review-World-Models: WAM-as-critic row (AtomVLA offline GRPO) - Review-In-Context-Imitation: ICLR-visual-reasoning + RoboSSM - Review-VLA-Memory: structured-vs-parametric memory row (GaussMemory/PROMPT vs TempoFit/RoboSSM) - Review-Multitask-VLA: VLA-RL / LAR-MoE / AtomVLA / MoE-humanoid - Review-Realtime-Execution: single-step head row (IMLE-VLA) - Review-Humanoid-VLA: IROS bimanual/whole-body trend (3DFA/EquiBim/ULTRA/CEER/MoE-VLA) Linked all from the IROS 2026 survey. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Sep 5, 2026
  • Escape pipe in all in-table wikilinks wiki-wide (202 links, 17 files) GitHub-wiki table cells read a wikilink's separator | as a column delimiter, splitting the cell and breaking the link. Escape to \| in every table-row wikilink (Home nav, RSS-2026-Papers, topic surveys). Prose wikilinks left as plain | (render correctly outside tables). Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 11, 2026
  • Weave ICML 2026 evidence into deep-dive surveys; revise three verdicts 14 State-of-the-Field sections gain ICML 2026 findings from the 99-paper index: recipes and latent-action supervision (VLANeXt, From-Pixels-to-Tokens, XR-1), MoT dual-systems and shortcut counters, the 9-paper efficiency cluster (Reflex 50Hz, GridS -76% FLOPs, XPU profile, latent reasoning -90%), reward/critic and model-based RL (VLAC, VLAW +39.2%), memory (HiMe/SOMA/CAPS), world models (DreamDojo 44kh, LAC-WM, dWorldEval), dexterous (DexMachina/DECO/Tabero/CTSRL), cross-embodiment (OXE-AugE, latent motion codes), evaluation (LIBERO-Gen, VLA-Arena, FixBench, TRAP). Verdicts revised: forgetting milder than assumed; discrete-token verdict scoped to robot-action auxiliaries; WM-evaluator action gap first crack. Home synced. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Decision map: prominent deep-dive links + uniform detail-page template Home fold-outs now lead with a heading-level "Deep dive ->" link and compress trend/approaches/limitations into a labeled 3-row table. The 12 detail-page State-of-the-Field sections are rewritten to one template (Verdict quote + Trend + Approaches-and-trade-offs + optional Established-findings + Limitations, dated Aug 2026); the three standalone surveys get matching headers with structure legends. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Survey-depth topic pages: 12 State-of-the-Field updates + 3 new surveys Each decision-map topic's detail page now carries a dated July-2026 survey section: trend arc through the latest venues, approach taxonomy with definitions and trade-offs, and current limitations. Three previously page-less topics get dedicated surveys: Human-Video Transfer (emergence/decoupling/synthesis fork + decision guide), VLA Evaluation (indictment + 2026 toolkit + emerging norms), Real-Time Execution (RTC->Legato arc + approach comparison). Home fold-outs link the full surveys; Reviews catalog updated. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 27, 2026
  • Integrate ICRA 2026 into cross-paper in-depth reviews Wove ICRA 2026 developments into 9 cross-paper reviews (drawing only from the already-verified ICRA topic/per-paper pages — no new web claims), mapping ICRA papers onto each review's existing taxonomy: - VLA-Architecture, Dexterous-Manipulation, RL, VLA-Memory, Goal-Image-Conditioning, System-0-1-2, Cross-Embodiment, VLM-Action-Connection, WAM-vs-VLA-Robustness Six of these had zero ICRA content before. 0 dangling wikilinks. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jun 1, 2026
  • Add ICRA 2026 (Vienna) VLA & manipulation analysis - New venue: ICRA hub + ICRA-2026 index + ICRA-2026 VLA & Manipulation Survey (5,088 submissions record; ~2,820 archival + 131 late-breaking; 11 themed clusters) - 10 web-verified per-paper pages: VLA-Reasoner, FD-VLA, OmniVLA(nav), Dexora, Flow Policy Optimization, MAP-VLA, Goal-VLA, LightVLA, VLA-Practicality, Galaxea+G0 - Cross-linked into Home, sidebar, and topic reviews (Architecture, RL, Memory, Goal-Image-Conditioning, Dexterous, AsyncVLA) - 0 dangling wikilinks across 212 pages Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jun 1, 2026
  • Source-verification audit: fix fabrications, rename PI slugs, link hygiene, add foundational pages - Re-verified all 195 pages vs original sources; removed 47 confirmed fabricated tables/numbers, restored 19 false-positive deletions (full-PDF re-check) - Renamed PI tech-report slugs ICLR-2026-pi07/pi06/RECAP -> PI-pi07/PI-pi06/PI-RECAP (these are PI technical reports, not ICLR 2026 papers); updated 171 wikilinks - Fixed 60 broken wikilinks -> 0 dangling across the wiki - Added foundational pages: OpenVLA, ReKep, AgiBot World Colosseo, RoboBrain 2.0; linked from Home + sidebar + venue indexes Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jun 1, 2026
  • Add in-depth cross-paper review of VLA memory architectures Covers 18 papers spanning π-series MEM, ICLR 2026 (MemoryVLA, HAMLET, Ctrl-World), CoRL 2025 (Long-VLA), and recent 2024-2026 preprints (MemER, SAM2Act+, ContextVLA, CronusVLA, ReMem-VLA, EchoVLA, ExpReS-VLA, Keyframe-Chaining VLA, MAP-VLA, TraceVLA, LoHoVLA, EvoVLA, RoboMemory). Page structure: - Why memory matters; HAMLET's 29.2 -> 76.4% headline delta on GR00T N1.5 - Six-category taxonomy (compressed video tokens, cross-attention memory banks, symbolic/cognitive memory, retrieval-based episodic stores, raw-frame history, planning-token hierarchies) - Per-paper deep-dives with arXiv links - Cross-axis comparison table (cheapest plug-in, most interpretable, longest horizon, best for spatial / on-device / hierarchical tasks) - Trends: from stateless (2023) to hybrid multi-scale (2026); plug-and-play on frozen backbones; retrieval eating robot memory - Open questions: selective write, hour-scale horizon, episodic vs parametric, interpretability probes, standard benchmark gap - Practical decision guide for shipping VLAs Cross-links added in HAMLET, MemoryVLA summary pages; sidebar and Home updated to surface the new review. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Apr 17, 2026