Doubt v0.5.0 — Agent Skills portability benchmark
Agent Skills portability: from documentation to evidence
Doubt v0.5.0 adds a six-source evidence map and a contribution-ready benchmark
for testing one unmodified Agent Skill across Claude Code, Codex, GitHub
Copilot, Cursor, and Gemini CLI.
What is new
- Interactive Agent Skills portability map
with 5 claims, 6 evidence items, 6 official sources, 1 contradiction, and 1
explicit unknown. - Frozen synthetic fixture plus direct, implicit, and negative prompt classes.
- Machine-readable protocol, JSON Schema, result template, and zero-dependency
result validator. - Exact client/configuration metadata, raw-output paths, artifact paths, and
SHA-256 receipts are required. - Honest
pass,fail, andblockedruns are accepted. - Single-client submissions are forbidden from claiming cross-client
behavioral equivalence.
The committed baseline is deliberately 0 submitted clients. Contribute the
first reproducible result in issue #9.
Install from npm
npx doubt-ai demo --out doubt-demo.htmlPackage: doubt-ai@0.5.0
Verification
- 34/34 automated tests
- 12/12 adversarial evidence-contract cases
- 5 topics / 25 keyed reader-study tasks
- Clean install and CLI smoke test from the public npm Registry
- CI on Node.js 18, 20, and 22
- GitHub Pages and the dogfooded Evidence Contract Action are green
Install the Agent Skill:
gh skill install alsoleg89/doubt doubtGate evidence maps in GitHub Actions:
- uses: alsoleg89/doubt@v0.5.0Release asset SHA-256:
2ad56697cd9f80a0853b9ff7f4766afc7b70c33ae784f000e1bc0cfb5ce99c0a