Repository navigation
v0.6.2-beta | Native CPU Embedding model support
Release Notes: Local CPU Embeddings & Sandbox Optimization Update
This update introduces local CPU embeddings, optimizes sandbox boot times for Linux hosts, and stabilizes memory extraction pipelines.
Architectural Changes: Local CPU Embeddings
- Integrated
llama-goC++ bindings to support lightning-fast, offline CPU inference for semantic search using GGUF models. - Portable Configuration Paths: Settings (such as
LocalEmbeddingModelPath) are now stored using a~/prefix substitution engine. This guarantees that a sharedconfig.jsonremains perfectly portable across different users, machines, and the internal Docker sandbox without "path poisoning." - Updated the Setup Script to natively prompt for and download compatible GGUF embedding models.
Sandbox & Host Parity
- Auto-Detect Binary Optimization: The
loa-sandboxboot sequence now dynamically inspects the host'sloaexecutable usingfileand ELF validation. If the host binary is compatible with the Debian container (e.g., Linux users), the sandbox imports the binary and skips compilation entirely for an instant boot. - Cross-OS Compatibility: For macOS and Windows users, the sandbox gracefully falls back to syncing source code and compiling the Linux binary inside the container, guaranteeing a working environment regardless of the host OS.
- Volume Mount Integrity: Fixed volume mounts to properly map the host's
.loadirectory directly to the container user's logical home, ensuringconfig.jsonportability. The host's global configuration is explicitly mounted Read-Only to prevent the sandbox from destructively overriding host settings.
Bug Fixes & Stability
- Memory Extraction: Fixed a structural bug in the memory extraction pipeline where parsing logic from message history arrays was failing or extracting incomplete contexts.
- UI Label Rendering: Repaired a CSS/JS rendering bug in the settings modal where configuration labels and dropdowns (specifically the Local Model Path) were displaying squished or hidden under certain viewport constraints.
Default Configuration Adjustments
To better align with modern 32B+ parameter models and deep architectural reasoning capabilities, default parameters have been upgraded:
- Max Plan Depth increased from
5to25. - Total Context Budget increased from
16,000to100,000.
Fix Github Workflow
- Fix github release build workflow