You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
This commit was created on GitHub.com and signed with GitHub’s verified signature.
[1.9.0] — 2026-06-12
Added
Model-family acceptance release gate: manifest-driven all-family strict acceptance for basics, discrete, continuous, hierarchical, multi-agent, precision, structured, gridworld, and scaling-study fixtures.
Cross-step evidence ledger: release ledger now links Step 3/5/6/11/12/15/16/23 statuses, artifact links, telemetry presence, renderer/execution status, and concrete skip reasons per family.
Interpretability summaries: per-family summaries now include variable/edge inventories, matrix-shape tables, telemetry presence, optional trace previews, renderer/execution status, and artifact links.
Changed
Continuous and hierarchical Step 11/12 outcomes are explicit profiled unsupported skips with concrete reasons, not raw render/execute failures accepted by profile math.
v1.7.0 is retired as a foundation-only track; unfinished runtime-depth ambitions move forward into v2+ reliability and orchestration milestones.
Current test evidence updated to 2,399 collected tests; final full-suite release evidence is recorded in TO-DO.md, README.md, and test documentation after the v1.9 release gate rerun.
Fixed
Removed the model-family acceptance reason-pattern fallback that could reclassify failed renderer/executor steps as unsupported success.
Hardened strict acceptance so profiled unsupported steps must be skipped before execution and failed Step 11/12 summaries fail closed.
Prevented cross-framework analysis from reading stale repo-tracked output/ artifacts during isolated /tmp acceptance runs.
Relaxed an environment performance smoke threshold to match other slow module smoke tests and avoid full-suite load false negatives.
Validation
PR #12 checks passed: Bandit, CodeQL, analyze (python), dependency-review, markdown-audit, security, test (3.11), test (3.12), test (3.13).
Local full suite: 2381 passed, 17 skipped, 1 xfailed with Ollama integration tests ignored.
Local collect-only: 2,399 collected tests with Ollama integration tests ignored.
All-family strict acceptance: passed for 9 manifest families; continuous and hierarchical Step 11/12 are explicit profiled unsupported skips with concrete reasons and no raw failed Step 11/12 counts.