v0.12.0
This release focuses on making adaptive routing more practical, observable, and reliable in real execution.
Highlights
- Improved adaptive routing so Sol can identify bounded delegation opportunities without becoming overly aggressive.
- Added safer support for single/sequential delegation when work is coupled but still separable.
- Kept unsafe parallelism blocked for mutable/shared-state workloads.
- Added routing rule/provenance telemetry for easier diagnosis.
- Hardened capability redaction across diagnostics, events, usage metadata, and legacy event paths.
- Fixed cancellation-state integrity across recovery, evidence scanning, semaphore waits, and activity projections.
- Improved Benchmark V3 campaign analysis and reproducibility metadata handling.
- Made benchmark history tests hermetic across clean CI environments.
- Updated documentation with the completed Benchmark V3 results.
Benchmark V3
Benchmark V3 completed 36/36 valid runs against the v0.11.0 baseline.
The campaign showed that v0.11.0 Adaptive Medium remained Solo across all tested workloads, adding orchestration overhead without using workers. Those findings directly informed the routing changes in v0.12.0.
The benchmark uses two repetitions and should be treated as directional evidence, not a statistically significant performance claim. v0.12.0 itself has not yet been compared against v0.11.0 in a new full benchmark campaign.
Validation
- 1,074 tests passing
- 0 failures
- 3 expected platform-specific skips
- Protocol smoke: PASS
- Benchmark fixture validation: PASS
npm audit: 0 vulnerabilities- Global packaged MCP smoke: PASS
- Real
gpt-5.6-lunadelegation verified through the packaged global MCP