Skip to content

Sirius 0.6 Beta 2

Pre-release
Pre-release

Choose a tag to compare

@tomislavrupic tomislavrupic released this 28 Aug 13:09
· 5 commits to main since this release

Sirius 0.6.0 Beta 2

Beta 2 promotes the verified Sirius 0.6 build used with Qwen3.8 Flash Next through MLX Serve on Apple Silicon.

What changed

  • Reliable switching between native MLX and MLX Serve model profiles.
  • Transactional rollback when a requested model cannot start.
  • Stronger download completeness checks for large model repositories.
  • Project constellations with compact neighboring stars and private or shared memory scope.
  • Streaming token-usage requests so direct model replies display measured decode speed.
  • Clearer local-runtime and memory-pressure diagnostics.
  • Graceful handoff between recorded inference backends.

Verified working path

  • Qwen3.8 Flash Next MLX Serve 4-bit loaded through Sirius.
  • Direct warm replies measured up to 77 tokens per second on the development Mac Studio.
  • fx workspace execution and optional macOS Control remain explicit opt-ins.
  • Model switching, local chat, project stars, and per-response speed display were exercised in the installed app.

Performance varies with hardware, prompt length, model state, tool use, and memory pressure. Agent-task timing is not the same as raw model decode speed.

Requirements

  • Apple Silicon Mac
  • macOS 14 or newer
  • Adequate storage and unified memory for the selected local model

Install

The beta is ad-hoc signed and is not notarized. Open the DMG, drag Sirius to Applications, then right-click Sirius and choose Open on first launch.

Package

  • Version: 0.6.0 Beta 2
  • Build: 27
  • File: Sirius-0.6.0-beta.2-arm64.dmg
  • SHA-256: 995b01ceb8cbd054e0b8259b0ca9d0d6249ec62451242ce76ead41e3b453ea06