Skip to content

v0.7.0 — cached API standard and full provider upgrade

Choose a tag to compare

@fire17 fire17 released this 31 Aug 07:16
· 2 commits to main since this release

What changed

  • Makes provider-native cached prompts the permanent localhost API standard: Anthropic cache_control + metadata and OpenAI prompt_cache_key/session routing are preserved end to end.
  • Adds seamless apiplan hotswap upgrade drain/replace on port 8787; the release cutover held 40/40 continuity probes.
  • Folds in the complete provider upgrade: Google Antigravity/Gemini, Ollama, media and video vision, credential single-flight/rotation recovery, tool-call fidelity, evidence-based health, and truncated-stream detection.
  • Addresses every API-capable model in the live Codex catalog, including gpt-reserve and codex-auto-review, while excluding models marked unsupported.

Verified

  • 272 tests pass, 0 fail, 808 assertions after cross-platform probe isolation.
  • All 7 performance budgets pass: 23 ms client startup, 3 ms owned dispatch+drain, 56 MB idle daemon.
  • Live cache receipts: 16,226 Anthropic cache-read tokens and 4,864 OpenAI cached tokens.
  • Live provider matrix: Claude Opus 5, GPT-5.6-Sol, Gemini 3.7 Flash and local heretic all returned the exact requested result.
  • 37 global commands installed; doctor all clear after command sync.