Sirius 0.6 Beta 2
Pre-release
Pre-release
·
5 commits
to main
since this release
Sirius 0.6.0 Beta 2
Beta 2 promotes the verified Sirius 0.6 build used with Qwen3.8 Flash Next through MLX Serve on Apple Silicon.
What changed
- Reliable switching between native MLX and MLX Serve model profiles.
- Transactional rollback when a requested model cannot start.
- Stronger download completeness checks for large model repositories.
- Project constellations with compact neighboring stars and private or shared memory scope.
- Streaming token-usage requests so direct model replies display measured decode speed.
- Clearer local-runtime and memory-pressure diagnostics.
- Graceful handoff between recorded inference backends.
Verified working path
- Qwen3.8 Flash Next MLX Serve 4-bit loaded through Sirius.
- Direct warm replies measured up to 77 tokens per second on the development Mac Studio.
- fx workspace execution and optional macOS Control remain explicit opt-ins.
- Model switching, local chat, project stars, and per-response speed display were exercised in the installed app.
Performance varies with hardware, prompt length, model state, tool use, and memory pressure. Agent-task timing is not the same as raw model decode speed.
Requirements
- Apple Silicon Mac
- macOS 14 or newer
- Adequate storage and unified memory for the selected local model
Install
The beta is ad-hoc signed and is not notarized. Open the DMG, drag Sirius to Applications, then right-click Sirius and choose Open on first launch.
Package
- Version: 0.6.0 Beta 2
- Build: 27
- File:
Sirius-0.6.0-beta.2-arm64.dmg - SHA-256:
995b01ceb8cbd054e0b8259b0ca9d0d6249ec62451242ce76ead41e3b453ea06