Escape pipe in all in-table wikilinks wiki-wide (202 links, 17 files)
GitHub-wiki table cells read a wikilink's separator | as a column
delimiter, splitting the cell and breaking the link. Escape to \| in
every table-row wikilink (Home nav, RSS-2026-Papers, topic surveys).
Prose wikilinks left as plain | (render correctly outside tables).
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Weave ICML 2026 evidence into deep-dive surveys; revise three verdicts
14 State-of-the-Field sections gain ICML 2026 findings from the
99-paper index: recipes and latent-action supervision (VLANeXt,
From-Pixels-to-Tokens, XR-1), MoT dual-systems and shortcut counters,
the 9-paper efficiency cluster (Reflex 50Hz, GridS -76% FLOPs, XPU
profile, latent reasoning -90%), reward/critic and model-based RL
(VLAC, VLAW +39.2%), memory (HiMe/SOMA/CAPS), world models (DreamDojo
44kh, LAC-WM, dWorldEval), dexterous (DexMachina/DECO/Tabero/CTSRL),
cross-embodiment (OXE-AugE, latent motion codes), evaluation
(LIBERO-Gen, VLA-Arena, FixBench, TRAP). Verdicts revised: forgetting
milder than assumed; discrete-token verdict scoped to robot-action
auxiliaries; WM-evaluator action gap first crack. Home synced.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Decision map: prominent deep-dive links + uniform detail-page template
Home fold-outs now lead with a heading-level "Deep dive ->" link and
compress trend/approaches/limitations into a labeled 3-row table.
The 12 detail-page State-of-the-Field sections are rewritten to one
template (Verdict quote + Trend + Approaches-and-trade-offs +
optional Established-findings + Limitations, dated Aug 2026); the
three standalone surveys get matching headers with structure legends.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Survey-depth topic pages: 12 State-of-the-Field updates + 3 new surveys
Each decision-map topic's detail page now carries a dated July-2026
survey section: trend arc through the latest venues, approach
taxonomy with definitions and trade-offs, and current limitations.
Three previously page-less topics get dedicated surveys:
Human-Video Transfer (emergence/decoupling/synthesis fork + decision
guide), VLA Evaluation (indictment + 2026 toolkit + emerging norms),
Real-Time Execution (RTC->Legato arc + approach comparison). Home
fold-outs link the full surveys; Reviews catalog updated.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Integrate ICRA 2026 into cross-paper in-depth reviews
Wove ICRA 2026 developments into 9 cross-paper reviews (drawing only from the
already-verified ICRA topic/per-paper pages — no new web claims), mapping ICRA
papers onto each review's existing taxonomy:
- VLA-Architecture, Dexterous-Manipulation, RL, VLA-Memory, Goal-Image-Conditioning,
System-0-1-2, Cross-Embodiment, VLM-Action-Connection, WAM-vs-VLA-Robustness
Six of these had zero ICRA content before. 0 dangling wikilinks.
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Add ICRA 2026 (Vienna) VLA & manipulation analysis
- New venue: ICRA hub + ICRA-2026 index + ICRA-2026 VLA & Manipulation Survey
(5,088 submissions record; ~2,820 archival + 131 late-breaking; 11 themed clusters)
- 10 web-verified per-paper pages: VLA-Reasoner, FD-VLA, OmniVLA(nav), Dexora,
Flow Policy Optimization, MAP-VLA, Goal-VLA, LightVLA, VLA-Practicality, Galaxea+G0
- Cross-linked into Home, sidebar, and topic reviews (Architecture, RL, Memory,
Goal-Image-Conditioning, Dexterous, AsyncVLA)
- 0 dangling wikilinks across 212 pages
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Source-verification audit: fix fabrications, rename PI slugs, link hygiene, add foundational pages
- Re-verified all 195 pages vs original sources; removed 47 confirmed fabricated
tables/numbers, restored 19 false-positive deletions (full-PDF re-check)
- Renamed PI tech-report slugs ICLR-2026-pi07/pi06/RECAP -> PI-pi07/PI-pi06/PI-RECAP
(these are PI technical reports, not ICLR 2026 papers); updated 171 wikilinks
- Fixed 60 broken wikilinks -> 0 dangling across the wiki
- Added foundational pages: OpenVLA, ReKep, AgiBot World Colosseo, RoboBrain 2.0;
linked from Home + sidebar + venue indexes
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Integrate NeurIPS 2025 papers into in-depth review pages
Review-VLA-Architecture.md:
- Add 13 NeurIPS 2025 papers to §3 paper table: Knowledge Insulation
(Spotlight), ChatVLA-2, DreamVLA, VLA-OS, Fast-in-Slow, ThinkAct,
Chain-of-Action, Real-Time Chunking, ReinFlow, VideoVLA, CogVLA,
BridgeVLA, DynaGuide.
- Extend category exemplar lists:
· B (Flow matching): Knowledge Insulation + RTC + ReinFlow
· E (World-model / VAM): DreamVLA + VideoVLA + SAMPO + OSVI-WM + RLVR-World
· F (Hierarchical / dual-system / MoE): Fast-in-Slow, ChatVLA-2,
ThinkAct, VLA-OS
· G (Reasoning / CoT): Chain-of-Action, ThinkAct, Robot-R1, DreamVLA
· I (Small / efficient): CogVLA
· J (Sensory-augmented): BridgeVLA (3D → 2D heatmap I/O)
- Update Trends 2 and 3 to highlight NeurIPS 2025 as the conference
that cemented dual-system (Trend 2) and deepened world-model-in-VLA
(Trend 3).
RL.md:
- Update header: now covers CoRL 2025 + NeurIPS 2025 + ICLR 2026.
- New §3a "Making flow-matching policies RL-trainable (log-prob-exact)":
ReinFlow + DSRL ancestor.
- Extend §7 (reasoning-token RL) with ThinkAct + Robot-R1.
- New §9 (Empirical studies & human-preference RL):
What-Can-RL-Bring (PPO > DPO/GRPO empirical) + APO (binary-signal
RLHF from HRI).
- Update reading order: insert What-Can-RL-Bring as step 2, ReinFlow
as step 5, ThinkAct as step 6.
Review-Cross-Embodiment.md:
- Add Grasp2Grasp (NeurIPS 2025, Princeton) to §3 paper table.
- Extend Category C (embodiment-invariant latents) with Grasp2Grasp
as a grasp-level Schrödinger-Bridge exemplar.
All three reviews now reflect that NeurIPS 2025 sits between CoRL 2025
and ICLR 2026 in the lineage — not just chronologically but substantively,
contributing key published formalisms (KI, RTC, ReinFlow) and algorithm
comparisons (What-Can-RL-Bring) that ICLR 2026 builds on.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Add cross-venue RL topic section + two new RL paper pages
New top-level topical landing:
- RL.md: cross-venue RL-for-VLA topic page with 8 recipe-based subsections
(residual RL, RL-token interface, outcome-conditioned, world-model RFT,
stage-aware reward shaping, VLM-as-reward, reasoning-token RL, scaling
infrastructure) plus adjacent test-time composition and evaluation
infrastructure. Includes mermaid recipe-decision diagram, comparison
tables, cross-cutting open questions, and a suggested reading order.
New paper pages:
- ICLR-2026-RL-Tokens: Physical Intelligence's RL Tokens (March 2026 tech
report). Tiny RL-token interface bolted onto frozen pi-0.6; tiny actor +
critic do online RL in minutes. Up to 3x speedup on screwdriving / zip-
tying / ethernet / power-cord; surpasses human teleop speed. Sits with
PLD/RFS as the production-engineering branch of residual RL.
- ICLR-2026-SimpleVLA-RL: ICLR 2026 paper (arXiv 2509.09674, PRIME-RL).
Scalable RL framework for VLA built on veRL with VLA-specific sampling /
parallelization / rendering / loss. SOTA LIBERO; beats pi-0 on RoboTwin
1.0 & 2.0; introduces the 'pushcut' phenomenon (RL discovers manipulation
patterns absent from SFT data).
Restructure:
- Home.md: now organized by two axes — Conferences and Topics. RL added as
the first topic; tree diagram updated to show paper pages reachable from
both axes.
- _Sidebar.md: new 'Research topics' section linking to RL; ICLR-2026 RL
group cross-references the topic page.
- ICLR-2026.md: RL section now points to cross-venue [[RL]] landing.
- ICLR-2026-VLA-Manipulation-Survey.md §5: links to [[RL]] topic landing
and adds RLT + SimpleVLA-RL rows to the recipe table.
- All RL paper pages: back-link footer now reads
'← Back to [[ICLR-2026]] · Topic: [[RL]]' so each paper is reachable from
both venue and topic axes.
Removed:
- ICLR-2026-RL-for-VLA.md (consolidated into top-level RL.md)
Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com>