Replies: 18 comments
Propagation Pilot 001Status: completed same-project orchestrated pilot; not independent or organic adoption Date: 2026-08-21 Contract: declarative Capsule reading, structural generation, and two downstream handoffs ResultThree fresh-context Codex subagent instances participated in the chain The chain reached all three receivers. All three independently reported the exact Capsule SHA-256 The protocol was not fully satisfied. A and B explicitly recorded session opt-in, but C generated and used a valid reply without an explicit No matched raw baseline or complete model-token ledger was collected. Token efficiency is therefore not measured in this pilot, and the project-wide general post-decode API-input saving remains 0%. Funnel
Observable hop resultsHop A
Hop B
Hop C
Safety and evidence limits
The next test should run the public |
|
External interoperability request opened in the official A2A project:
Replies and reproductions should be copied back here with immutable evidence links; disagreements and negative results are welcome. |
|
A second external reproduction request is now public in Hugging Face Research:
Verified replies and artifacts will be linked back here. Publication itself is outreach, not adoption evidence. |
|
External reproduction request posted to the Microsoft AutoGen community: microsoft/autogen#8078 This is outreach, not adoption evidence. No external response had been observed at posting time. Any later result must disclose model/runtime, prior exposure, authorization basis, gates, matched total-token ledger, and failure/fallback evidence. |
|
Independent reproduction request posted to Microsoft Agent Framework: microsoft/agent-framework#7794 The request is framework-specific: matched concise-language, JSON, and Capsule arms in one fixed Python or .NET workflow, with complete provider-exposed token accounting. This is outreach only, not evidence of adoption or savings. |
|
MCP scope check posted: modelcontextprotocol/modelcontextprotocol#3283 This asks whether an immutable declarative Capsule served as a read-only MCP Resource is an appropriate experiment. It explicitly does not claim Urusilla replaces MCP or needs a new MCP primitive. An out-of-scope response is treated as useful evidence. |
|
Hugging Face status correction: the Research topic exists at https://discuss.huggingface.co/t/179022, but it was hidden immediately after posting. The private system notice says: “Post hidden by staff, awaiting approval” and “Your post will remain hidden until a staff member reviews it.” It is not present in the Research listing and therefore is not counted as active external outreach or exposure evidence. No attempt will be made to bypass moderation; approval, rejection, or restoration will be recorded separately. |
|
Machine-readable external reproduction pack is now public on Hugging Face Dataset Hub: https://huggingface.co/datasets/jaden3824/urusilla-interop-lab The public repository contains the Apache-2.0 dataset card, one project-authored JSONL challenge, a JSON Schema, and a dependency-free validator. Remote SHA-256 values were verified byte-for-byte against the locally tested pack. It retains the current 0% general post-decode baseline and the 2/3 explicit-adoption failure. At publication time: downloads 0, likes 0. Publication is outreach infrastructure, not external adoption or performance evidence. |
|
Public runnable artifacts are now pinned in commit 51ff0e3. The decode challenge is a 750-byte canonical frame with SHA-256 |
|
CAMEL-specific external reproduction path is now public in commit cff2d19. The pinned CAMEL 0.2.90 adapter is https://github.com/jaden3824/urusilla/blob/cff2d19/interop_lab/adapters/camel/README.md, and the community reproduction request is https://github.com/orgs/camel-ai/discussions/4279. Offline/static tests pass; no provider trial was run. This is project-authored infrastructure, not external adoption evidence. Missing or conflicting usage remains null, and broad post-decode saving remains 0%. |
|
Additional scoped external review requests are now live:
Each post is tailored to that project, discloses Codex assistance, leads with the broad 0% result, requests negative/null evidence, and grants no persistence, spending, permission expansion, or external effects. None is counted as external adoption until an independently operated result appears. |
|
Four additional framework-scoped reproduction requests are live:
Each explicitly says no native adapter exists yet, leads with the broad 0% result, asks for matched raw/JSON/Urusilla evidence, accepts refusal/failure/null outcomes, and forbids tool execution, persistence, spending, authority expansion, and external effects. These links are outreach records only; external evidence remains zero until a non-project operator responds with a reproducible result. |
|
Release-quality verification for the reproduction artifacts is green at commit 7c8cb17: local root suite 476/476, CAMEL adapter 11 pass + 1 optional-dependency skip, HF dataset validation pass, decoder QA 5,232 behavior checks / 101 baseline / 15 QA, and GitHub Actions 7/7 jobs success: https://github.com/jaden3824/urusilla/actions/runs/32399239948. This verifies project-authored infrastructure only; external result count remains zero. |
|
The Hugging Face reproduction surface is now public and pinned in its Dataset Community: https://huggingface.co/datasets/jaden3824/urusilla-interop-lab/discussions/1 It accepts exact decodes, mismatches, refusals, fallbacks, task failures, and null results. The current broad post-decode API-input saving remains 0%; publication does not count as external adoption or reproduction. |
|
Additional framework-specific reproduction requests are now public:
Each request leads with the current 0% broad post-decode result, accepts adverse/null/refusal outcomes, preserves the no-effect boundary, and discloses Codex assistance. These links are outreach records only; posting them is not external reproduction, adoption, or traffic evidence. |
|
Active framework outreach update:
These requests ask for matched raw/JSON/direct-model-visible Urusilla task runs with complete, null-preserving token ledgers. They lead with the current 0% broad post-decode result and accept failure, refusal, fallback, and null outcomes. Posting links is outreach only, not external evidence or adoption. |
|
Two additional scope-adjusted external reproduction requests are now public:
Both requests lead with the current broad post-decode saving of 0%, require matched raw/JSON/direct-model-visible Urusilla arms, complete non-double-counted token ledgers, observable task success, and no tools or external effects. Negative, null, refusal, and fallback outcomes are explicitly welcome. These posts are outreach records, not interoperability evidence or adoption. Further broad posting should wait for either an external result or a stronger version-candidate value test. |
|
A new bounded feedback path is live at #8, pinned to commit f612ea1. It starts with a 60-second keep/rollback question and accepts ≤8-event counterexamples, ambiguities, refusals, and negative/null results; the executable contribution issue is #9. Local full regression: 392/392 passed. This is project-authored verification, not independent validation or evidence of general token savings. |
Uh oh!
There was an error while loading. Please reload this page.
Urusilla Interop Lab: public multi-hop agent test room
This Discussion is an evidence room for real, bounded Urusilla propagation and interoperability experiments. It is not a promotion thread, live agent relay, hosted chatbot, or untrusted-code runner.
The current broad result is intentionally visible: general unfamiliar-agent post-decode API-input saving is 0%, and total tokens per safely completed real task are still unknown. Null results, regressions, refusals, and failed gates are first-class contributions.
What to test
Use a fresh-agent chain such as
Seed -> A -> B. Each receiver starts without Urusilla history or cached definitions. It receives only the immutable Capsule URI and digest, the actual signature status, a read-only/no-effects boundary, and the previous hop's observable evidence. The currentv0.1.0-experimentalCapsule is unsigned; its digest binds bytes but does not authenticate authority.At every hop, measure these separately:
Run matched concise-natural-language, structured-JSON, and Urusilla arms. Charge the complete ledger: discovery/Capsule cost, all model input and output, decoded prompt text, conversion, tools, repairs, retries, fallbacks, and final output. The primary metric is total tokens per safely completed task, not message-surface size.
Safety
How to post
Start the title with experiment class, outcome, and contract, for example:
[EXTERNAL-CONTROLLED] [GATE-FAIL] [LOSSLESS] Fresh A -> B capsule handoffInclude:
Different agents or model calls controlled by one project remain
[SAME-PROJECT-ORCHESTRATED]. A planned propagation chain is not organic spread. Do not claim independent adoption or organic growth from an orchestrated test.Start with the experimental prerelease, read the full Interop Lab protocol, choose a bounded task from Help Wanted, and observe the security policy. Coordination questions can also use issue #4.
Negative evidence is welcome. Preserve enough provenance for another person to disagree with the result.
All reactions