Skip to content

History / Review Multitask VLA

Revisions

  • IROS 2026 major papers: upgrade to in-depth with paper figures + arXiv links; add VLA-RL page - Added arXiv links + embedded the papers' own architecture/results figures (with attribution) to: AtomVLA (2603.08519, 2-stage WM-critic pipeline), 3D FlowMatch Actor (2508.11002, 3D-scene-token architecture; +21d->16h / 0.5->18.2Hz), RoboSSM (2509.19658, results: SSM holds vs ICRT collapse as demos->32, 16x longer prompts), TempoFit (2603.07647, layer-wise KV-memory method figure). - New IROS-2026-VLA-RL page (2505.18719, Tsinghua/NTU): trajectory-as- conversation online RL + VLM process-reward; OpenVLA-7B +4.5% LIBERO, real 60->90%, inference-scaling; framework figure embedded. Linked from survey and Multi-Task-VLA review. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Sep 7, 2026
  • IROS 2026: add 5 full-paper analyses + reflect into existing in-depth reviews Full-paper pages (abstract-verified from official program): - IROS-2026-AtomVLA: subtask-aware VLA + latent-WM scoring for offline GRPO (WAM lens) - IROS-2026-3D-FlowMatch-Actor: CMU/NVIDIA unified single/dual-arm 3D policy, +41.4% PerAct2, ~30x faster (bimanual SOTA) - IROS-2026-EquiBim: symmetry-equivariant bimanual policy - IROS-2026-IMLE-VLA: single-step cIMLE action head, 55Hz, LIBERO 98.0% (efficiency) - IROS-2026-ICLR-Visual-Reasoning: in-context imitation with image-space reasoning traces Reflected IROS 2026 into existing reviews: - Review-World-Models: WAM-as-critic row (AtomVLA offline GRPO) - Review-In-Context-Imitation: ICLR-visual-reasoning + RoboSSM - Review-VLA-Memory: structured-vs-parametric memory row (GaussMemory/PROMPT vs TempoFit/RoboSSM) - Review-Multitask-VLA: VLA-RL / LAR-MoE / AtomVLA / MoE-humanoid - Review-Realtime-Execution: single-step head row (IMLE-VLA) - Review-Humanoid-VLA: IROS bimanual/whole-body trend (3DFA/EquiBim/ULTRA/CEER/MoE-VLA) Linked all from the IROS 2026 survey. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Sep 5, 2026
  • Add in-depth Review pages for HiMoE-VLA and DyGRO-VLA (multi-task VLA cluster) - Review-HiMoE-VLA: hierarchical depth-wise MoE (shallow AS-MoE per action space, deeper HB-MoE per embodiment, central dense consolidation; AS-Reg/ HB-Reg; 32 experts top-4). Negative-transfer ablation (dense pi0 -0.259 vs HiMoE +0.186), LIBERO 97.8 / real xArm7 75.0 / ALOHA 63.7. Framed as the MoE-routing cluster of multi-task fixes. - Review-DyGRO-VLA: cross-task RL scaling — protect a shared latent (offline info-theoretic pretrain) then optimize grouped RL residuals (a=Δa+a_base, K-critic ensemble, entropy load-balance). LIBERO 92.7->97.1, Long 85.2->95.0. - Repointed Multi-Task VLA links to the new Review pages; added in-depth pointers to the ICLR/ICML venue entries; catalog entries in Reviews. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 14, 2026
  • Add Multi-Task VLA review + MergeVLA page - Review-Multitask-VLA: why one VLA fails across many tasks (6 failure modes: negative transfer/gradient conflict, non-mergeability, multi-task conflict, catastrophic forgetting, routing confusion, instruction collapse) + a 7-cluster solution landscape (merging, MoE, gradient control, skill decomposition, instruction grounding, continual, adapters) with a decision guide and open questions. Anchored on MergeVLA (CVPR 2026). - CVPR-2026-MergeVLA: per-paper page (non-mergeability diagnosis + task-masked LoRA / cross-attention-only action expert / test-time task router). - Cross-linked from Reviews catalog and Home topic reviews. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 14, 2026