Desktop Dev Build #196.1 (67553e2)
Pre-release
Pre-release
Automated desktop build artifacts for commit 67553e2fdc03803e07ea6e0ab4043dece06cfa15.
Workflow run: https://github.com/LettuceAI/app/actions/runs/31650317025
Changes since previous Desktop build desktop-dev-195-1-5029153
Compare: desktop-dev-195-1-5029153...67553e2
67553e2docs(changelog): add the 2.2.0 release entry41c4e9bfeat(group-chats): add per-character model and system prompt overrides22b5267fix: align companion timelines with the session clock4e49f5fi18n: translate the model duplicate and safetensors format stringse97bf0bfeat(image-gen): accept safetensors in the sd.cpp model selector and repo search0b8d722fix(widgets): hide companion-only widgets in roleplay and group chats117cd58fix(chats): show total session message counts in widgets356a975feat(models): add a duplicate action to the model menu37d473efix(navigation): preserve settings return path through playgrounda679cf5fix(images): honor model storage and playground exit145915ffix(images): stop sd.cpp when Lettuce exits04e7b6ffix(prompts): drop the unused scene protocol predicate052e9ebfix(images): treat unset scene generation as disabled5c4e2a8fix(prompts): gate each scene image protocol on its own model kindb69837cfix(storage): repair group chat schema drift9b34d2afeat(images): support literouter and custom openai-format image providers90b97aefeat(llama): measure the device compute reserve instead of guessing ita328ebbfix(llama): report kv for offloaded blocks only, add a real-model plan probe2f2d672fix(llama): derive the mtp draft reserve from real weights and kvb688f7ffix(llama): reserve attention scratch for query heads not kv heads766f8eafix(llama): distribute multi-gpu layers using real per-unit costsfb20eddfix(llama): subtract resident weights when recommending a vram context89e603ffix(llama): size kv from per-layer geometry including sliding windowsa03be64style(llama): drop inline comments from the offload changes072ada8fix(llama): derive kv size from declared head dims and ggml block sizesdc2fe98fix(llama): price gpu offload from real per-unit tensor sizes47f5c75fix(llama): load and fit bundled mtp tensors for embedded mtp models1257c8dchore(llama): bump llama-cpp-rs to b10327 and opt into mtp tensor loadinga6c99cbfix(ollama): normalize stray system messages to satisfy strict templates (#85)04b97c7fix(ci): require Ada and Blackwell CUDA kernels1a97db2fix(companion): preserve shared memory when creating chatsd444236fix(nanogpt): show exhausted quota as fully used04c8e48feat(companion): preserve canonical message timelineb7f2deffeat(companion): add evidence-based continuity across chat episodes3e1ef0cfeat(companion): preserve soul and relationship continuity across chats4350526fix(companion): make message timestamps unmistakably system-authored53e8426feat(sync): replace legacy replication with transactional sync v2 (#84)6a5eca7fix(sync): reject mismatched builds and repair stale state