Sirius 0.6.1 Beta 1
Pre-release
Pre-release
·
2 commits
to main
since this release
Sirius 0.6.1 Beta 1
This is a debug-oriented follow-up to 0.6.0 Beta 2. It makes the production
provider boundary explicit so MLX Serve can be tested independently from the
experimental Native MLX and Elastic paths.
What changed
- MLX Serve is the default inference backend for Direct Qwen, fx, and BrowserCode.
- Native MLX is available only with the explicit development opt-in
SIRIUS_NATIVE_INFERENCE=1. - The status strip reports the selected backend (
MLX SERVE,NATIVE MLX, or
ELASTIC) instead of presenting ECO as if it were an engine choice. - Model residency remains lazy so opening Sirius does not load a large model.
- Release metadata, update detection, and the public landing-page download now
target this 0.6.1 Beta 1 artifact. - Verified updates can now be installed with a single
RESTART & INSTALL
action: Sirius hands the SHA-256-checked DMG to a short-lived helper, exits,
preserves a rollback copy, verifies the replacement, and relaunches.
Debug focus
When testing this build, record the backend label, selected model, available
unified memory, loopback health on port 11234, and direct-model tokens/second
before evaluating fx latency. The Flash-Next 4-bit preview remains an explicit
96 GB-class model selection and can fail its memory preflight on a busy Mac.
Package
- Version: 0.6.1 Beta 1
- Build: 28
- File:
Sirius-0.6.1-beta.1-arm64.dmg