Skip to content

Doubt v0.5.0 — Agent Skills portability benchmark

Choose a tag to compare

@alsoleg89 alsoleg89 released this 31 Jul 07:57
· 22 commits to main since this release

Agent Skills portability: from documentation to evidence

Doubt v0.5.0 adds a six-source evidence map and a contribution-ready benchmark
for testing one unmodified Agent Skill across Claude Code, Codex, GitHub
Copilot, Cursor, and Gemini CLI.

What is new

  • Interactive Agent Skills portability map
    with 5 claims, 6 evidence items, 6 official sources, 1 contradiction, and 1
    explicit unknown.
  • Frozen synthetic fixture plus direct, implicit, and negative prompt classes.
  • Machine-readable protocol, JSON Schema, result template, and zero-dependency
    result validator.
  • Exact client/configuration metadata, raw-output paths, artifact paths, and
    SHA-256 receipts are required.
  • Honest pass, fail, and blocked runs are accepted.
  • Single-client submissions are forbidden from claiming cross-client
    behavioral equivalence.

The committed baseline is deliberately 0 submitted clients. Contribute the
first reproducible result in issue #9.

Install from npm

npx doubt-ai demo --out doubt-demo.html

Package: doubt-ai@0.5.0

Verification

  • 34/34 automated tests
  • 12/12 adversarial evidence-contract cases
  • 5 topics / 25 keyed reader-study tasks
  • Clean install and CLI smoke test from the public npm Registry
  • CI on Node.js 18, 20, and 22
  • GitHub Pages and the dogfooded Evidence Contract Action are green

Install the Agent Skill:

gh skill install alsoleg89/doubt doubt

Gate evidence maps in GitHub Actions:

- uses: alsoleg89/doubt@v0.5.0

Release asset SHA-256:

2ad56697cd9f80a0853b9ff7f4766afc7b70c33ae784f000e1bc0cfb5ce99c0a