Skip to content

History / Changelog

Revisions

  • Add DYNA-2 in-depth review (Dyna Robotics World-Action Model launch) Company announcement (Aug 10, 2026), not a paper: WAM on ~1M h human egocentric video with no robot data in pre-training, joint next-frame+ next-action, claimed first human-to-robot scaling law smooth over 1k->1M h (~50x EgoScale), 87% vs 46% zero-shot over DYNA-1. Reviewed with an explicit vendor-claim caveat (no technical paper/benchmark/ weights). Filed under Latest Papers; cross-linked from World-Models, Human-Video-Transfer, Reviews. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 11, 2026
  • Add Latest Papers tracker + omega-0 and Stellar VLA in-depth reviews New Latest-Papers.md preprint tracker (pre-publication reviews) and two figure-illustrated in-depth reviews: omega-0 (arXiv 2608.06375, whole- body humanoid latent-predictive World Action Model; 81.8% on 11 household tasks vs 44.5% psi-0; ships 40h omega-HOME dataset) and Stellar VLA (arXiv 2511.18085, continual imitation learning with a Dirichlet-Process knowledge space + knowledge-routed MoE, 1% replay). Cross-linked from Home, Reviews, sidebar, Humanoid-VLA, World-Models. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 11, 2026
  • Add DreamZero in-depth review (World Action Models are Zero-shot Policies) NVIDIA's 14B video-diffusion World Action Model (arXiv 2602.15922): jointly predicts video+action, >2x over SOTA VLAs on unseen-env/ unseen-object real-robot evals, 38x inference stack (DreamZero-Flash) for 7 Hz closed-loop control, video-only cross-embodiment transfer. Fig. 4 architecture embedded. Cross-linked from World Models review, Home lab-programs, Reviews catalog, and sidebar. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 10, 2026
  • Fix RSS survey render timeout: split paper index to RSS-2026-Papers The survey exceeded GitHub wiki's render budget (262 wikilinks/40KB). Full session tables (116 linked rows) move to a dedicated RSS-2026-Papers index page; the survey keeps a pointer and drops to ~147 links / 21KB. Hub and changelog note the split. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Complete RSS 2026 per-paper coverage; add themed-list insights 101 new reference pages (verbatim verified abstracts + metadata + topic-review links) bring all 116 in-scope RSS papers to full page coverage. Survey updated: all session tables linked, condensed block expanded to five full tables with abstract glosses, bolded themed- list mentions linked (100), and each of the 8 themed reading lists now opens with a technology-level insight paragraph. RSS hub updated. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Weave ICML 2026 evidence into deep-dive surveys; revise three verdicts 14 State-of-the-Field sections gain ICML 2026 findings from the 99-paper index: recipes and latent-action supervision (VLANeXt, From-Pixels-to-Tokens, XR-1), MoT dual-systems and shortcut counters, the 9-paper efficiency cluster (Reflex 50Hz, GridS -76% FLOPs, XPU profile, latent reasoning -90%), reward/critic and model-based RL (VLAC, VLAW +39.2%), memory (HiMe/SOMA/CAPS), world models (DreamDojo 44kh, LAC-WM, dWorldEval), dexterous (DexMachina/DECO/Tabero/CTSRL), cross-embodiment (OXE-AugE, latent motion codes), evaluation (LIBERO-Gen, VLA-Arena, FixBench, TRAP). Verdicts revised: forgetting milder than assumed; discrete-token verdict scoped to robot-action auxiliaries; WM-evaluator action gap first crack. Home synced. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Decision map: prominent deep-dive links + uniform detail-page template Home fold-outs now lead with a heading-level "Deep dive ->" link and compress trend/approaches/limitations into a labeled 3-row table. The 12 detail-page State-of-the-Field sections are rewritten to one template (Verdict quote + Trend + Approaches-and-trade-offs + optional Established-findings + Limitations, dated Aug 2026); the three standalone surveys get matching headers with structure legends. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Aug 5, 2026
  • Survey-depth topic pages: 12 State-of-the-Field updates + 3 new surveys Each decision-map topic's detail page now carries a dated July-2026 survey section: trend arc through the latest venues, approach taxonomy with definitions and trade-offs, and current limitations. Three previously page-less topics get dedicated surveys: Human-Video Transfer (emergence/decoupling/synthesis fork + decision guide), VLA Evaluation (indictment + 2026 toolkit + emerging norms), Real-Time Execution (RTC->Legato arc + approach comparison). Home fold-outs link the full surveys; Reviews catalog updated. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 27, 2026
  • Rewrite Home decision map as an insight map Each of the 15 topics is now a collapsible entry carrying the trend arc through the latest venues, competing approaches with definitions and trade-offs, and current limitations - replacing the plain link table. All claims use wiki-verified numbers (OAT/Legato, RECAP, LDA-1B/mimic-video, VLM4VLA +18.1, DexGrasp-Zero/OHRA, Psi-0, LIBERO-X pyramid, LBM verdicts, camera-frame EEF scaling law). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 27, 2026
  • Redesign Home: themed decision map, stat strip; move rules to Maintenance Decision map split into four themed tables (Building / Running & improving / Data & evaluation / Embodiment) with a current-answer column; stat strip and emoji headers added; reviews section as a compact table; Page Format / Maintenance Rule section moved to the new Maintenance page, linked from the footer. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Restructure navigation: top-level sidebar + Reviews catalog page New Reviews.md holds the complete in-depth-review catalog (topic reviews, lab programs, per-paper long-forms, RSS 2026 figure pages). Sidebar slimmed to top-level only: reviews hub + six star topics, model lineages, ML hub, one link per venue year (venue pages already index their papers), foundational refs. Home merges its two review sections into one compact section pointing at the catalog. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Knowledge Graph: reflect RSS 2026 across all views + new cluster Overall map gains RSS 2026 (and missing ICML) venue nodes; reviews graph gains Tactile-VLA/Humanoid-VLA nodes with RSS-labeled edges; lineages extended (RECAP marked RSS oral, RTC->Legato branch, Psi-0 adoption, new TRI-LBM and LIBERO robustness lines); world-model cluster adds mimic-video, LDA-1B (unified-WM branch), Qwen-RobotWorld; new fifth view maps the RSS 2026 improvement-loop cluster (6 threads -> 15 clickable pages). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Refresh Home research decision map with RSS 2026 findings Four new question rows (improvement-from-experience, human-video-vs- robot-data fork, co-training data selection, credible evaluation) and five rows updated with RSS evidence (Legato/OAT, LDA-1B/mimic-video, ViTacFormer/CGP, cross-hand transfer, Psi-0). Header stats and start-here pointer updated to the RSS 2026 survey. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Expand remaining RSS 2026 pages with figures; add RSS to sidebar Figure-illustrated reviews with confirmed arXiv IDs for OAT, Legato, Contact-Grounded Policy, DexGrasp-Zero, One Hand to Rule Them All, and LIBERO-X (affiliations, awards, and L1-L5 benchmark details added). Sidebar: RSS entry in Conferences, a full RSS 2026 section, and the recent Qwen-series + Psi-0 in-depth reviews. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • RSS 2026 major papers: figure-illustrated 1-page reviews Extract and embed key figures from the arXiv originals of the nine major RSS 2026 papers (pi*0.6/RECAP, Psi-0, LDA-1B, H2R-Emergence, mimic-video, LBM co-training study, ViTacFormer, PolaRiS, HoMMI), each with an explanatory caption; confirm arXiv IDs and add hardware/ backbone details learned from the papers (SharpaWave hands, CVAE architecture, Cosmos-Predict2 backbone, LDA-1B per-domain numbers). Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Add RSS 2026 conference survey, Psi-0 in-depth, and 14 per-paper pages RSS 2026 (Sydney, Jul 13-17): 210 accepted papers parsed from the official program, ~116 manipulation/hand/humanoid papers in scope, all abstracts verified. Survey covers six threads: RL-from-experience (pi*0.6/RECAP flagship), human-video transfer (emergence vs decoupling), video/world models vs VLA backbones, contact-as-representation, cross-embodiment dexterous hands, and evaluation infrastructure. New: RSS venue hub, Review-Psi0 (full-paper in-depth: 800h human video + 30h robot data beats 10x corpora by >40pp on Unitree G1), per-paper pages for LDA-1B, H2R-Emergence, mimic-video, LBM co-training study, ViTacFormer, DexGrasp-Zero, One-Hand, Contact-Grounded Policy, PolaRiS, LIBERO-X, OAT, Legato, HoMMI. PI-RECAP updated with RSS camera-ready results. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 25, 2026
  • Qwen-VLA review: re-verify against arXiv v2 (Jun 1) v1->v2 adds a no-T2A baseline (60.9%; T2A = +10.2 pp) and clarifies that T2A shares the downstream action representation (chunk-first-frame delta EEF). T2A ablation figures aligned to v2 one-decimal precision; all benchmark tables unchanged between versions. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 24, 2026
  • VLM4VLA review: re-verify against arXiv v2, fix pi0 Calvin Task-3 value v1->v2 diff contains a single substantive change: pi0 Calvin Task-3 0.786 -> 0.686 (fixes internal sum; total 3.509 unchanged). Header now records version history and the Tsinghua x Qwen affiliation split. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 24, 2026
  • Add in-depth Qwen-RobotNav and Qwen-RobotWorld reviews; update program page RobotNav (2606.18112): parameterized observation interface, 15.6M corpus, agentic EQA SOTA, supersedes Qwen-VLA nav by ~15pp. RobotWorld (2606.17030): 20B double-stream MMDiT, frozen Qwen2.5-VL action encoder, EWK 8.6M corpus, Scene2Robot; 1st on EWMBench/DreamGen. Program page: suite rows upgraded to deep-read, backbone/lambda doctrine qualified, System-2 slot marked demonstrated, watch-list items 4-5 updated; sibling cross-links added. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 24, 2026
  • Add cross-paper review: Qwen Team's VLA Program VLM4VLA -> Qwen-VLA -> Qwen-Robot Suite (RobotManip/RobotNav/RobotWorld): shared doctrine (Qwen3.5-4B, flow matching, lambda=0.1 VL co-training, synthetic-data scaling, language-as-interface), diagnostic-to-flagship trace, internal contradictions between the two flagship VLAs, and lab-program positioning vs PI/GR00T/TRI/Gemini. Cross-links from the three constituent reviews; indexed in Home/Changelog. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 24, 2026
  • Add in-depth Qwen-RobotManip review (arXiv 2606.17846) Alignment-first scaling thesis, camera-frame delta EEF + CaPE, 38,100h open-data corpus with 24,808h human-to-robot synthesis, RoboTwin-IF/XE benchmarks, RoboChallenge Table30-v1 generalist #1. Index in Home/Changelog; cross-link from Qwen-VLA review. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jul 24, 2026
  • Add in-depth review: Demystifying Action Space (EEF vs Joint, absolute vs delta) Long-form review of arXiv 2602.23408 (ICML 2026): the action-abstraction taxonomy (joint vs task/EEF space; absolute vs delta; chunk-wise vs step-wise), the EEF-vs-joint verdict (joint=stability & scales; EEF=generalization/transfer), the O(k)-vs-O(1) noise-propagation result behind chunk-wise delta, and the paper's practical guidelines. Embeds paper Figures 1/2/3/5 (attributed). Linked from the one-pager, sidebar, Home deep-dives, and Changelog. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jun 11, 2026
  • Restructure Home around a Research Decision Map; split out Changelog New front-page structure: scale/freshness header + start-here, (1) intent-based Research Decision Map, (2) Latest Updates (moved up, dated, links to Changelog), (3) Core Topic Reviews (cross-paper only, one-liners restored), (4) merged venue survey table, (5) per-paper/series deep-dives, (6) foundational refs, (7) page-format/maintenance rule. Adds the new VLA Training Frameworks and RoboMME pages; de-dups OpenVLA; splits a dated Changelog.md (sidebar-linked). Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>

    @Heungwoo Heungwoo committed Jun 11, 2026