Skip to content

v0.0.8

Latest

Choose a tag to compare

@wsauret wsauret released this 03 Sep 13:49

This release strengthens account-aware usage, paid-capacity controls, session recovery, workflow consistency, and learned-skill maintenance. It also improves terminal visibility for startup, notifications, usage, diffs, and agent-launched work while reducing runtime memory, repeated event folding, and diagnostic growth

Added

  • Show profile accounts, exact quota windows, monthly and range-based spend, credits, automatic billing status, pricing provenance, and count-only metrics backed by a canonical request ledger that attributes requests and credential rotations to the correct account lineage
  • Require explicit per-account consent before using paid subscription capacity, notify when review is needed, expose consistent controls after login and at safe boundaries, and bind admission to current profile usage, credits, prepaid capacity, and automatic billing state
  • Add customizable subscription usage alert thresholds and deliver provider, runner, runtime, and usage warnings as durable project-scoped toasts and notifications
  • Add exact-provider model controls, reviewer subtype overrides, role-specific default models, and a dedicated designer model seat for interface judgment
  • Discover code through linked directories while keeping Flywheel-owned directories private
  • Add safe child-worktree execution with project and execution-directory authority enforced across nested agents and tools
  • Add reusable standalone agent methods for research, root-cause analysis, verification, UAT, implementation, and specialist reviews, with elegance-first synthesis and composable target-specific guidance
  • Add interface-planning and interface-review capabilities through model-visible skills and a design reviewer that evaluates hierarchy, theming, states, spacing, motion, and copy
  • Make the advisor available to workflow subagent steps and allow explorers to inspect GitHub diffs with gh
  • Add learned-skill health tooling with flywheel skills doctor, deterministic moved-path repair, body-budget reporting, development-only pointer checks, and explicit guidance for moving oversized content into references
  • Add learned-skill maintenance that surveys library structure, plans merges and repairs, applies validated units atomically with --apply, preserves unchanged files without repeating them, and continues after isolated refusals
  • Add learned-skill consolidation with dry-run support, recoverable archives, semantic auditing, automated cleanup after bundled-skill changes, and a CLI option to force cleanup
  • Add durable interrupted-turn reporting that records which tools ran and which files changed so a resumed agent can continue without repeating completed work
  • Pin workflow definitions and handoff contracts when runs are admitted so active and resumed runs retain the exact revision they started with
  • Show blocking startup phases and migration progress before the runtime is ready, and surface worker startup failures instead of leaving an empty loading screen
  • Show all edited files in compact transcript rows and independently preview each file
  • Organize settings and provider controls into expandable groups
  • Add independently scrollable, easier-to-scan subscription account cards while hiding ChatGPT Codex limits by default
  • Add /doctor for checking learned-skill health from a session

Fixed

  • Prevent duplicate OpenAI tool calls from corrupting sessions by deduplicating completed response items with their stable call identity
  • Avoid false provider outages after Escape, system sleep, local runtime failures, silent stream closure, or unsupported inference by requiring positive wire evidence before durable suppression and allowing planning to select a suppressed sole provider as a last resort
  • Preserve session state across provider resets and keep degraded or malformed-history sessions available for explicit restart
  • Make crash recovery scoped and lossless by bounding recovery to the current conversation, durably preserving queued input, distinguishing possibly unexecuted tools, and recording deliberate turn dismissal
  • Keep stalled recovery within one logical turn, retry stalled responses in smaller work chunks, resume capped turns from durable context without duplicate work, and keep provider retry messages out of model context
  • Let Escape interrupt model retries
  • Restore live session-list updates in every window by attaching projection streaming to the first protocol client rather than the first internal listener
  • Keep agent-launched workflows visible by isolating malformed session rows, retrying failed catalog reads, and projecting execution directories from available metadata
  • Keep background notifications durable across session reopen and deliver them through chat, workflow, and other runner types
  • Enforce paid-capacity decisions against exact, fresh evidence, preserve the newest review state, reject stale writes, honor zero caps, and keep credit failover consent-safe across restarts
  • Keep paid-capacity review available for verified accounts when another account cannot be verified, while continuing to fail closed when no evidence is verifiable
  • Correctly label unavailable paid sources as off rather than declined, explain unfunded or disabled sources, and offer review only when it can authorize spending
  • Show paid-capacity save failures and preserve in-progress decisions while reviewing
  • Keep profile switching atomic and isolated across roles, usage, refresh generations, model labels, and downstream admission while adopting already prepared runtime state
  • Validate saved model targets against the current provider account and preserve provider catalog caches across upgrades, treating undecodable old caches as absent until refreshed
  • Restore SuperGrok device login
  • Continue automatically after quota failover and refresh token-based finite Copilot quotas
  • Keep complete repository instructions in model context, including nested imports from the repository root and safely resolved symlinked instructions
  • Improve agent prompting and model-visible guidance to prefer empirical design evidence, resolve uncertainty before implementation, highlight recommended implementation routes, require explicit workflow approval, and verify delivered code against approved intent
  • Improve agent prompting and model-visible review guidance to preserve verified constraints, choose one minimal remedy, make low-priority findings opt-in, and avoid unsupported best-practice complexity
  • Change agent prompting and model-visible skill guidance to correct contradicted claims rather than append beneath them, preserve like sections together, state exact body-budget remedies, and avoid compressing new content merely to fit
  • Allow the advisor to perform read-only shell inspection
  • Keep managed command output streaming, and update agent guidance so Bash timeouts are reserved for commands with real deadlines
  • Bound native worker generations and failed admissions so stalled replacements cannot grow worker populations indefinitely
  • Prevent runtime replacement and shutdown from overlapping, stop role updates after shutdown, and release client sockets exactly once so idle cleanup can run
  • Keep the idle reaper alive after stalled or rejected work checks by treating uncertainty as live work, recording the failure, and rearming the check
  • Create required event-store companion tables before migrations so older stores no longer crash startup
  • Prevent unreadable or abandoned in-flight recovery state from blocking activation or resurrecting deleted sessions, and refuse file mutations when their recovery state cannot be prepared
  • Keep paused workflows retryable after pruning, remove superseded facts without deleting current per-step state, and preserve pinned workflow revisions
  • Keep model overrides reactive during active workflows
  • Require explicit user approval before launching workflows
  • Reject impossible learned-skill operation recovery states, keep maintenance manifests strict, renew cleanup claims during long model requests, and make cleanup leases and retries crash-safe
  • Prevent learned-skill audits from following instructions inside the packages they inspect and pin the admitted audit identity
  • Honor configured skill archive windows and recheck recent activity before startup cleanup archives a skill
  • Continue learned-skill maintenance after a refused unit, while stopping on request-wide or partial-write failures and reporting usable outcomes
  • Accept dragged file paths in edit operations
  • Keep mutation previews complete while redacting recognized credential values only, allowing ordinary public path references to remain visible
  • Show syntax-highlighted file-write previews and previews for large diffs, while keeping retained full diffs collapsed until expanded
  • Show and edit complete Ask User responses, including preserving typed Other answers while navigating choices
  • Keep tab actions correct after terminal resize
  • Keep provider account details responsive and show an error when a modal cannot load
  • Keep session role labels current
  • Render Luna shell calls as Bash
  • Remove misleading process controls from transcript rows
  • Make subagent failures explain how to recover and preserve reviewer effort after reload
  • Make configuration notices actionable
  • Show update alerts as a dismissible prompt badge
  • Restore quota, startup, and usage progress bars with compact modal sizing
  • Classify unknown POSIX directory entries instead of silently omitting them
  • Make local installation and updates work with standard Windows permissions
  • Search and list through Windows junctions by shipping linked-directory support and comparing canonical paths in the same Windows path form, while still blocking junctions into Flywheel-owned directories

Changed

  • Use space-separated CLI subcommands and --project-root; colon-delimited subcommands, bare maintain, and --base-dir are no longer accepted
  • Count credit-backed requests as spend and distinguish prepaid capacity from automatic billing
  • Show Copilot profile caps only when usage is pooled
  • Keep passive notifications separate from Escape so interrupting a turn does not dismiss them
  • Cap editable skill bodies at 12,000 characters, allow existing oversized bodies only to shrink, and direct case-specific material into references
  • Move routine per-call diagnostics to debug, gate every record family by a declared tier, and roll diagnostic files at a byte limit
  • Store transient operation recovery checkpoints beside the append-only history so repeated renewals and provisional updates replace one live value instead of accumulating durable facts
  • Prune superseded chat provisionals and workflow view rows when sessions close, with explicit maintenance commands available for historical stores

Performance

  • Reuse registered event projections and incrementally replay only new facts instead of rebuilding full aggregate history on repeated reads, reducing conversation-tree folds during activation
  • Bound retained projections per family, limiting heavy conversation and tool projections to eight and reducing measured retained conversation state from 316.8 MB to 18.6 MB
  • Fold background workflow queues incrementally with one held projection per session instead of replaying every queue from the beginning on each pass
  • Reuse held session metadata projections and read only requested session IDs instead of rebuilding and sorting the full catalog
  • Add flywheel events maintain --refresh-stats-snapshot and include it in full maintenance, reducing a measured cold workspace-stats fold from 508 ms to 4 ms
  • Remove per-fragment durable writes from the streaming turn path and recover interrupted work from already committed tool and file facts
  • Prevent unbounded diagnostic files and routine INFO traffic from filling disks and causing SQLite write failures