Releases: eric861129/Clean-Code-AI-Collaboration-Skill
Release list
v0.5.2:Profile 證據修復與 M2 Pilot
v0.5.2:Profile 查閱證據與比較組隔離
本版修正評測協定並公開全新 M2 Pilot 證據,保留失敗與證據不足的結論。正式發布不代表所有 Profile 效果通過,也不代表 Stable 或跨模型驗證。
本版內容
- 將 Profile 查閱內容分成必讀、允許與禁止集合,避免合法的額外 Core 查閱遭誤判,並維持跨 Profile 污染檢查。
- 新增版本化診斷、Oracle 未執行原因及固定執行來源,強化 Archive Bytes 與 Canary Freeze 綁定。
- 公開 34 組全新 Subject Runs 與 34 份匿名 Review;首批 11 組 Canary 通過協定 Gate,未混入正式結果。
- 更新正式安裝版本為 v0.5.2;歷史 Result 不改寫,四份 Profile 的技術指引語意未調校。
實際結果與限制
| Profile | Outcome | 成熟度 | 解讀 |
|---|---|---|---|
| C# | no_difference | Experimental | 固定直接比較未觀察到增益 |
| Python | inconclusive | Experimental | Core Only Comparator 的 Hash Claim 不完整,仍有證據缺口 |
| TypeScript | failed | Experimental | Treatment 的驗收失敗與評分退步保留於結果 |
| React | passed | Beta | 本輪固定比較支持舊請求狀態覆寫防護的改善,不能推論所有 React 任務 |
34 組正式執行皆為單一 Attempt,31 組 passed、3 組 automatic_failure,沒有 Timeout、基礎設施重試或選擇性重跑。Go、Rust、Java、Vue 仍為 Planned,不可 Routing 或 Packaging。本版未執行 Full Run、Stable Promotion 或 Specialist Package 公開發布。
Subject 請求設定為 gpt-5.6-sol/high,Reviewer 為 gpt-5.6-terra/max;實際服務模型、Token 與網路隔離 Telemetry 未取得,不據此宣稱額外認證。
固定來源與公開證據
- 正式發布 Commit:
5facc8c79f82a21aabeaa98d9f7eef85400b3e81。 - 受測 Skill 固定為 v0.5.1/
4ce3a697eb7eea36ecddd937ffa915f931e9524a,與發布 Suite v0.5.2 分開紀錄。 - 執行 Harness 固定為
bbdfb0618dc65b0a56ead0e79d0929c1acc305f8。 - Fixture:profile-pilot-v2/
133e2f739e5d943e853f4c9106d8965ecfac3018。 - 公開 Result SHA-256:
df733b4e0f65d76a0f6027f07f277e1b491bc975cb3a8ea8da2b1289b28122fd。 - 正式 Manifest 與 Benchmark 說明。
公開 JSON 僅做 CRLF → LF 與末尾換行正規化,解析後內容一致;原始 Bytes、轉換前後 SHA 與限制見 Benchmark。v0.5.1 歷史結果原樣保留。
驗證與安裝
本地 251 項測試完成,249 項通過、2 項 Windows Symlink 測試跳過。Compile、Profile Registry、Generated Documentation、Release-mode C# 範例包及 Agent Skill 格式驗證通過。
完整檢查亦由 開發分支 CI、main 雙平台 CI 與 Tag 雙平台 CI 核對,main 與 Tag 指向相同發布 SHA。CI 有 Action Node.js Runtime 棄用警告,本版驗證通過;Runtime 升級留待後續維護。公開證據可用 python -m unittest tests.test_profile_pilot_v052_replay -v 重播,需要完整 Git 歷史及公開 Fixture 存取,不需重新執行 Subject 或取得私有 Workspace。
安裝請依 README,切到 v0.5.2 Tag 後,只安裝 clean-code-ai-collaboration 資料夾。
Clean Code AI Collaboration Skill v0.5.1
v0.5.1 完成 C#、Python、TypeScript 與 React 的 M2 Profile Pilot,公開 34 組執行結果與匿名審查,並讓 Profile Metadata 可核對公開實證。此版本沒有將任何 Profile 升級為 Beta。
本次變更
- 評測 Harness 改由 Manifest 定義 Scenario 與比較組,保留 v0.3.0 的 36 組歷史契約。
- 固定八個 Scenario、34 個 Subject Runs 與
profile-pilot-v1Fixture Tag。 - 加入 Profile Outcome 計算、匿名 Review、Result Schema 與 Metadata Evidence 核對。
- 補上不完整 Result 的拒絕驗證,以及凍結歷史來源與公開 M2 結果的 CI 重播檢查。
- 公開 Oracle 命令移除本機絕對路徑;安裝文件固定為
v0.5.1。
實證結果與限制
- 34/34 組執行已進入 Terminal State:17
passed、17automatic_failure;17 組自動失敗的記錄原因皆為invalid_claim,沒有選擇性重跑。 - 八位隔離 Reviewer 完成 34 份匿名審查。
- C#、Python、TypeScript、React 的 Outcome 皆為
inconclusive,維持experimental,benchmark_status為pilot_recorded。 - 此結果尚不能支持四個 Profile 的增量效果;完整保留失敗與證據缺口。
- 這是固定模型、Client、Prompt 與 Fixture 下的一次 Pilot Repetition。Subject 使用固定的
v0.5.0Skill 來源;Suite 發布版本與受測來源分別記錄。 - Desktop Token Telemetry 與 Network Enforcement 資訊不可取得,均如實記錄,未推估。
- Go、Rust、Java、Vue 仍為
planned。本次沒有 Full Run、Stable Promotion 或 Specialist Package 公開發布。
可核對來源
- 公開 M2 Result
- Pilot Manifest
- Fixture Tag
- Result SHA-256:
12637bf780a156fb3589c810d0c6deeae8ce3abba9d2a54b8a5a06f5aeecadc9 - 凍結執行 Harness 來源:
c7c49df0b5a9b9a1bc49d6e36030c50f0fea372f;公開路徑去識別修正:983ca2f。CI 以歷史來源核對執行契約,避免把發布後的 Harness 當成原始執行版本。 - 歷史 Fixture Hash 依當時 Windows CRLF 換行與路徑排序重建;跨平台 CI 使用相同歷史條件核對,不改寫凍結值。重播需要完整 Git 歷史及固定 GitHub Fixture Archive 的網路存取。
驗證
- 本地完整測試:208 項,206 通過、2 項因 Windows 符號連結權限跳過。
- Compile、Profile Registry、Generated Documentation、Core 與 C# Sample Agent Skill 驗證均通過。
- Standards Review 與 Spec Review 各找到一項驗證缺口,均已修復並加入回歸測試。
- 合併 commit 的 GitHub Actions:Ubuntu、Windows 全部通過,commit 為
4ce3a697eb7eea36ecddd937ffa915f931e9524a。
安裝
依 README 固定安裝 v0.5.1;Skill Frontmatter 版本為 0.5.1。
Clean Code AI Collaboration Skill v0.5.0
v0.5.0 重點
這個版本把單一 Core Skill 升級成可依 Changed Module 組合技術 Profile 的架構。Core 仍是授權、風險路徑、驗證與輸出契約的唯一來源;Profile 只補充語言與 Framework 的技術語意。
Composable Profiles
- 新增 C#、Python、TypeScript 與 React 四個 Experimental Profile。
- TypeScript 與 React 可依 Changed Module 同時套用,並依 Language、Framework 順序載入。
- Go、Rust、Java 與 Vue 以 Planned 狀態保留,不會被 Routing 或 Packaging 誤用。
- 支援 Explicit Profile 與 Core Only;Explicit Profile 仍須通過 Availability、Applicability 與 Conflict 檢查。
Registry、Routing 與文件
- 新增 Profile Metadata、Catalog 與 JSON Schema Draft 2020-12 契約。
- 新增 deterministic Changed Module Routing、Reason Code 與 Polyglot Fixtures。
- Runtime 使用 Generated Profile Index,不要求 Consumer Agent 解析 YAML。
- 新增 Profile Authoring Guide、生命週期、Evidence 與 Reviewer Checklist。
Specialist Packaging
- 新增可重現、Fail-closed 的 Specialist Packager 與獨立 Package Manifest Schema。
- Release mode 全程從 Git Blob 讀取輸入,避免 Dirty Checkout 與 EOL 影響產物。
- Output Directory 已存在、Path Escape、Symlink Escape、Profile Conflict 等情況都會停止,不覆寫既有內容。
- 本次沒有附加 Specialist ZIP;
dist/clean-code-ai-csharp/仍是 Git ignored 的本機 Sample。
驗證
- 175 項 Repository Tests 通過。
- Ubuntu 與 Windows GitHub Actions 都通過 Full Tests、Compile、Profile Validator、Generated Drift Check、Release-mode C# Sample Build,以及 Core/Sample Open Standard Validation。
- v0.4.0 的已發布 Benchmark Evidence 保持不變,歷史重播改從已發布 Tag 的 Git Blob 驗證。
成熟度限制
四個新 Profile 目前都是 Experimental,只代表文件、Metadata、Routing 與 Contract Tests 已完成。尚未執行 M2 Pilot,不宣稱 Beta/Stable,也不宣稱跨 Repository 的普遍品質提升。
安裝
git clone https://github.com/eric861129/Clean-Code-AI-Collaboration-Skill.git
cd Clean-Code-AI-Collaboration-Skill
git checkout v0.5.0真正需要安裝的是 clean-code-ai-collaboration 資料夾。完整安裝方式、設定與 Profile 使用方式請閱讀 README。
Clean Code AI Collaboration Skill v0.4.0
v0.4.0 重點
這個版本讓使用者可以明確選擇 AI Agent 的開發節奏與驗證範圍,同時保留 Clean Code、Repository 規則、必要 Gate 與授權邊界。
新增開發節奏
auto:依風險、測試 Oracle、回饋速度、工作樹與授權判斷。direct:適合小型、可逆且行為已知的修改。tdd:先以 RED 說清楚需求缺口,再完成最小 GREEN。tcr:只在測試快速可靠、工作樹隔離且具備當次 Commit/Revert 授權時使用。characterization-first:先固定 Legacy Code 的現有可觀察行為。
新增驗證範圍
autofocusedrepositoryacceptance-e2emutation-assisted
開發節奏與驗證方式彼此獨立,例如可以使用 TDD 開發,再加入 E2E 驗證跨邊界結果。
安全與操作邊界
- 明確策略缺少前提時,Agent 必須停止並回報,不能偷偷改用其他節奏。
development_rhythm: tcr不等於授權 Commit 或 Revert。- 使用者選擇的驗證範圍不能取消 Repository 既有的必要 Gate。
- Codex adapter 維持 explicit-only,請明確指定
$clean-code-ai-collaboration。
README 與評測
README 新增快速開始、Prompt/AGENTS.md 設定方式、巢狀規則、TDD/TCR 範例、常見組合、FAQ,以及 Windows、macOS、Linux 的安裝與重播指令。
固定策略案例的無 Skill 對照為 5/9 符合決策契約;載入 v0.4.0 後,18 次執行為 18/18 通過。這項結果只支持固定情境下的策略契約解讀較一致,不衡量程式碼品質,也不宣稱一定節省 Token、時間或費用。
評測樹雜湊採用固定的 POSIX 相對路徑排序,已在 Windows 本機與 GitHub Actions Linux Runner 驗證可重播。
安裝
git clone https://github.com/eric861129/Clean-Code-AI-Collaboration-Skill.git
cd Clean-Code-AI-Collaboration-Skill
git checkout v0.4.0真正需要安裝的是 clean-code-ai-collaboration 資料夾。完整安裝方式與設定範例請閱讀 README。
Cross-language Benchmark Protocol v4 Pilot Freeze
This prerelease publishes the Desktop Subject Protocol v4 harness and the frozen twelve-run cross-language Pilot contract for clean-code-ai-collaboration v0.3.0.
Included
- Controller-owned immutable report fields instead of trusting Subject-provided hashes and baseline claims.
- Verifiable staged Skill inspection using the required
SKILL.mdpath and SHA-256. - Fixed TypeScript/React and Python/FastAPI manifest using Fixture tag
cross-language-v3. - Twelve valid Pilot terminal states, all passed.
- Sixty public, preservation, and acceptance Oracle commands, all succeeded.
- Anonymous packet review with no reported identity leakage or Freeze blocker.
- Frozen contract SHA-256:
8f932545c0cc1c642688c02b6bd319230ae1d3bd1bf0351efac8169fd599ad29.
Claim boundary
This is a Pilot contract freeze, not the final cross-language benchmark result. The remaining twenty-four Full Run sessions have not been executed, no aggregate winner is published, and Token/Tool Call/network telemetry remains not_available. These observations must not be generalized to every language, model, client, or repository.
Clean Code AI Collaboration Skill v0.3.0
Highlights
- Adds Lightweight, Standard, and Full Audit paths with one-way risk escalation.
- Separates delivery maturity into Prototype and Production-Ready without weakening authorization or risk gates.
- Routes validation through repository-native Formatter, Linter, Static Analysis, Build, Test, Security, Coverage, and Mutation gates according to the actual risk.
- Splits review output into a concise Standard contract and a traceable Full Audit contract.
- Narrows the activation scope and sets the Codex adapter to explicit-only with allow_implicit_invocation: false.
- Documents external-validity expansion across legacy .NET, TypeScript and React, Python, Java and Spring, repository instructions, test quality, models, clients, and blind review.
Evidence boundary
The published benchmark numbers remain the fixed v0.2.0 results. v0.3.0 introduces routing, delivery-readiness, toolchain, and documentation changes, but does not claim new cross-language success rates or Token savings.
Verification
- 38 repository contract and benchmark tests passed.
- Agent Skills validator passed.
- SKILL.md body remains within the 500-word limit at 490 words.
- Existing benchmark manifests and result files were not rewritten.