Add DYNA-2 in-depth review (Dyna Robotics World-Action Model launch)
Company announcement (Aug 10, 2026), not a paper: WAM on ~1M h human
egocentric video with no robot data in pre-training, joint next-frame+
next-action, claimed first human-to-robot scaling law smooth over
1k->1M h (~50x EgoScale), 87% vs 46% zero-shot over DYNA-1. Reviewed
with an explicit vendor-claim caveat (no technical paper/benchmark/
weights). Filed under Latest Papers; cross-linked from World-Models,
Human-Video-Transfer, Reviews.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add Latest Papers tracker + omega-0 and Stellar VLA in-depth reviews
New Latest-Papers.md preprint tracker (pre-publication reviews) and two
figure-illustrated in-depth reviews: omega-0 (arXiv 2608.06375, whole-
body humanoid latent-predictive World Action Model; 81.8% on 11
household tasks vs 44.5% psi-0; ships 40h omega-HOME dataset) and
Stellar VLA (arXiv 2511.18085, continual imitation learning with a
Dirichlet-Process knowledge space + knowledge-routed MoE, 1% replay).
Cross-linked from Home, Reviews, sidebar, Humanoid-VLA, World-Models.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add DreamZero in-depth review (World Action Models are Zero-shot Policies)
NVIDIA's 14B video-diffusion World Action Model (arXiv 2602.15922):
jointly predicts video+action, >2x over SOTA VLAs on unseen-env/
unseen-object real-robot evals, 38x inference stack (DreamZero-Flash)
for 7 Hz closed-loop control, video-only cross-embodiment transfer.
Fig. 4 architecture embedded. Cross-linked from World Models review,
Home lab-programs, Reviews catalog, and sidebar.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Fix RSS survey render timeout: split paper index to RSS-2026-Papers
The survey exceeded GitHub wiki's render budget (262 wikilinks/40KB).
Full session tables (116 linked rows) move to a dedicated
RSS-2026-Papers index page; the survey keeps a pointer and drops to
~147 links / 21KB. Hub and changelog note the split.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Complete RSS 2026 per-paper coverage; add themed-list insights
101 new reference pages (verbatim verified abstracts + metadata +
topic-review links) bring all 116 in-scope RSS papers to full page
coverage. Survey updated: all session tables linked, condensed block
expanded to five full tables with abstract glosses, bolded themed-
list mentions linked (100), and each of the 8 themed reading lists
now opens with a technology-level insight paragraph. RSS hub updated.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Weave ICML 2026 evidence into deep-dive surveys; revise three verdicts
14 State-of-the-Field sections gain ICML 2026 findings from the
99-paper index: recipes and latent-action supervision (VLANeXt,
From-Pixels-to-Tokens, XR-1), MoT dual-systems and shortcut counters,
the 9-paper efficiency cluster (Reflex 50Hz, GridS -76% FLOPs, XPU
profile, latent reasoning -90%), reward/critic and model-based RL
(VLAC, VLAW +39.2%), memory (HiMe/SOMA/CAPS), world models (DreamDojo
44kh, LAC-WM, dWorldEval), dexterous (DexMachina/DECO/Tabero/CTSRL),
cross-embodiment (OXE-AugE, latent motion codes), evaluation
(LIBERO-Gen, VLA-Arena, FixBench, TRAP). Verdicts revised: forgetting
milder than assumed; discrete-token verdict scoped to robot-action
auxiliaries; WM-evaluator action gap first crack. Home synced.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Decision map: prominent deep-dive links + uniform detail-page template
Home fold-outs now lead with a heading-level "Deep dive ->" link and
compress trend/approaches/limitations into a labeled 3-row table.
The 12 detail-page State-of-the-Field sections are rewritten to one
template (Verdict quote + Trend + Approaches-and-trade-offs +
optional Established-findings + Limitations, dated Aug 2026); the
three standalone surveys get matching headers with structure legends.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Survey-depth topic pages: 12 State-of-the-Field updates + 3 new surveys
Each decision-map topic's detail page now carries a dated July-2026
survey section: trend arc through the latest venues, approach
taxonomy with definitions and trade-offs, and current limitations.
Three previously page-less topics get dedicated surveys:
Human-Video Transfer (emergence/decoupling/synthesis fork + decision
guide), VLA Evaluation (indictment + 2026 toolkit + emerging norms),
Real-Time Execution (RTC->Legato arc + approach comparison). Home
fold-outs link the full surveys; Reviews catalog updated.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Rewrite Home decision map as an insight map
Each of the 15 topics is now a collapsible entry carrying the trend
arc through the latest venues, competing approaches with definitions
and trade-offs, and current limitations - replacing the plain link
table. All claims use wiki-verified numbers (OAT/Legato, RECAP,
LDA-1B/mimic-video, VLM4VLA +18.1, DexGrasp-Zero/OHRA, Psi-0,
LIBERO-X pyramid, LBM verdicts, camera-frame EEF scaling law).
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Redesign Home: themed decision map, stat strip; move rules to Maintenance
Decision map split into four themed tables (Building / Running &
improving / Data & evaluation / Embodiment) with a current-answer
column; stat strip and emoji headers added; reviews section as a
compact table; Page Format / Maintenance Rule section moved to the
new Maintenance page, linked from the footer.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Restructure navigation: top-level sidebar + Reviews catalog page
New Reviews.md holds the complete in-depth-review catalog (topic
reviews, lab programs, per-paper long-forms, RSS 2026 figure pages).
Sidebar slimmed to top-level only: reviews hub + six star topics,
model lineages, ML hub, one link per venue year (venue pages already
index their papers), foundational refs. Home merges its two review
sections into one compact section pointing at the catalog.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Knowledge Graph: reflect RSS 2026 across all views + new cluster
Overall map gains RSS 2026 (and missing ICML) venue nodes; reviews
graph gains Tactile-VLA/Humanoid-VLA nodes with RSS-labeled edges;
lineages extended (RECAP marked RSS oral, RTC->Legato branch, Psi-0
adoption, new TRI-LBM and LIBERO robustness lines); world-model
cluster adds mimic-video, LDA-1B (unified-WM branch), Qwen-RobotWorld;
new fifth view maps the RSS 2026 improvement-loop cluster (6 threads
-> 15 clickable pages).
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Refresh Home research decision map with RSS 2026 findings
Four new question rows (improvement-from-experience, human-video-vs-
robot-data fork, co-training data selection, credible evaluation) and
five rows updated with RSS evidence (Legato/OAT, LDA-1B/mimic-video,
ViTacFormer/CGP, cross-hand transfer, Psi-0). Header stats and
start-here pointer updated to the RSS 2026 survey.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Expand remaining RSS 2026 pages with figures; add RSS to sidebar
Figure-illustrated reviews with confirmed arXiv IDs for OAT, Legato,
Contact-Grounded Policy, DexGrasp-Zero, One Hand to Rule Them All,
and LIBERO-X (affiliations, awards, and L1-L5 benchmark details
added). Sidebar: RSS entry in Conferences, a full RSS 2026 section,
and the recent Qwen-series + Psi-0 in-depth reviews.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
RSS 2026 major papers: figure-illustrated 1-page reviews
Extract and embed key figures from the arXiv originals of the nine
major RSS 2026 papers (pi*0.6/RECAP, Psi-0, LDA-1B, H2R-Emergence,
mimic-video, LBM co-training study, ViTacFormer, PolaRiS, HoMMI),
each with an explanatory caption; confirm arXiv IDs and add hardware/
backbone details learned from the papers (SharpaWave hands, CVAE
architecture, Cosmos-Predict2 backbone, LDA-1B per-domain numbers).
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add RSS 2026 conference survey, Psi-0 in-depth, and 14 per-paper pages
RSS 2026 (Sydney, Jul 13-17): 210 accepted papers parsed from the official
program, ~116 manipulation/hand/humanoid papers in scope, all abstracts
verified. Survey covers six threads: RL-from-experience (pi*0.6/RECAP
flagship), human-video transfer (emergence vs decoupling), video/world
models vs VLA backbones, contact-as-representation, cross-embodiment
dexterous hands, and evaluation infrastructure. New: RSS venue hub,
Review-Psi0 (full-paper in-depth: 800h human video + 30h robot data
beats 10x corpora by >40pp on Unitree G1), per-paper pages for LDA-1B,
H2R-Emergence, mimic-video, LBM co-training study, ViTacFormer,
DexGrasp-Zero, One-Hand, Contact-Grounded Policy, PolaRiS, LIBERO-X,
OAT, Legato, HoMMI. PI-RECAP updated with RSS camera-ready results.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Qwen-VLA review: re-verify against arXiv v2 (Jun 1)
v1->v2 adds a no-T2A baseline (60.9%; T2A = +10.2 pp) and clarifies that
T2A shares the downstream action representation (chunk-first-frame delta
EEF). T2A ablation figures aligned to v2 one-decimal precision; all
benchmark tables unchanged between versions.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
VLM4VLA review: re-verify against arXiv v2, fix pi0 Calvin Task-3 value
v1->v2 diff contains a single substantive change: pi0 Calvin Task-3
0.786 -> 0.686 (fixes internal sum; total 3.509 unchanged). Header now
records version history and the Tsinghua x Qwen affiliation split.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add in-depth Qwen-RobotNav and Qwen-RobotWorld reviews; update program page
RobotNav (2606.18112): parameterized observation interface, 15.6M corpus,
agentic EQA SOTA, supersedes Qwen-VLA nav by ~15pp. RobotWorld (2606.17030):
20B double-stream MMDiT, frozen Qwen2.5-VL action encoder, EWK 8.6M corpus,
Scene2Robot; 1st on EWMBench/DreamGen. Program page: suite rows upgraded to
deep-read, backbone/lambda doctrine qualified, System-2 slot marked
demonstrated, watch-list items 4-5 updated; sibling cross-links added.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add cross-paper review: Qwen Team's VLA Program
VLM4VLA -> Qwen-VLA -> Qwen-Robot Suite (RobotManip/RobotNav/RobotWorld):
shared doctrine (Qwen3.5-4B, flow matching, lambda=0.1 VL co-training,
synthetic-data scaling, language-as-interface), diagnostic-to-flagship
trace, internal contradictions between the two flagship VLAs, and
lab-program positioning vs PI/GR00T/TRI/Gemini. Cross-links from the
three constituent reviews; indexed in Home/Changelog.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add in-depth Qwen-RobotManip review (arXiv 2606.17846)
Alignment-first scaling thesis, camera-frame delta EEF + CaPE,
38,100h open-data corpus with 24,808h human-to-robot synthesis,
RoboTwin-IF/XE benchmarks, RoboChallenge Table30-v1 generalist #1.
Index in Home/Changelog; cross-link from Qwen-VLA review.
Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
Add in-depth review: Demystifying Action Space (EEF vs Joint, absolute vs delta)
Long-form review of arXiv 2602.23408 (ICML 2026): the action-abstraction
taxonomy (joint vs task/EEF space; absolute vs delta; chunk-wise vs step-wise),
the EEF-vs-joint verdict (joint=stability & scales; EEF=generalization/transfer),
the O(k)-vs-O(1) noise-propagation result behind chunk-wise delta, and the
paper's practical guidelines. Embeds paper Figures 1/2/3/5 (attributed).
Linked from the one-pager, sidebar, Home deep-dives, and Changelog.
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Restructure Home around a Research Decision Map; split out Changelog
New front-page structure: scale/freshness header + start-here, (1) intent-based
Research Decision Map, (2) Latest Updates (moved up, dated, links to Changelog),
(3) Core Topic Reviews (cross-paper only, one-liners restored), (4) merged
venue survey table, (5) per-paper/series deep-dives, (6) foundational refs,
(7) page-format/maintenance rule. Adds the new VLA Training Frameworks and
RoboMME pages; de-dups OpenVLA; splits a dated Changelog.md (sidebar-linked).
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>