You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
This commit was created on GitHub.com and signed with GitHub’s verified signature.
[0.6.0] - 2026-09-08
Added
Agent self-improvement Quick Starts can now evaluate real target-AI responses through the selected model profile and retain response evidence as a run artifact.
The runner SDK now provides a credential-safe target-AI completion helper for evaluation assets.
The README includes a ten-second Quick Start guide, representative templates, and a response-level prompt-improvement example.
Changed
The managed prompt sample is intentionally minimal so the response-evaluation Quick Start can demonstrate evidence-backed improvement from a real behavior gap.
Fixed
Evaluation-run history totals now retain proposed and adopted improvements and reported issues from every supervisor iteration, rather than showing only the latest iteration.
Evaluation-run translations now apply only to the active Supervisor AI iteration or currently filtered result records, with independent loading state for each view.