Skip to content

v1.9.0

Choose a tag to compare

@github-actions github-actions released this 29 Aug 06:31
40ebec0

1.9.0 (2026-08-28)

Sumika v1.9.0 focuses on greater control, broader file support, and a smoother chat experience:

  • Reasoning controls: Users can select the model’s reasoning effort to balance speed and answer quality.
  • Improved skills: Skills can be referenced directly in messages and discovered across multiple configured locations.
  • DOCX attachments: Word documents can now be attached and used as chat context.
  • Better transcript interaction: Links and newly attached images are clickable, with improved readability and scrolling behavior.
  • Performance and reliability: Transcript scrolling uses less CPU, images are correctly oriented for vision models.

Features

  • add debug-gated MLX memory tracing (9fda3d0), closes #225
  • add first-valid multi-root skill discovery (a064355)
  • add selectable reasoning effort levels (935d73e)
  • add support for skill mentions in user messages (c2d0c36)
  • centralize model generation settings (f9fbb2b)
  • enhance content classification for CSS and TypeScript, add new test cases (0b8e82c)
  • implement link opening functionality in transcript views (56e25dd)
  • set Qwen 3.8 reasoning effort to medium (ea22cf1)
  • support DOCX chat attachments (#228) (98b3c22)

Bug Fixes

  • add viewport width change handling and update streaming revision test for scroller styles (fde9901)
  • correct formatting of FluidAudio package URL (c32613f)
  • enhance download progress reporting in model downloader tests (019385c)
  • enhance scroll view behavior and add viewport width handling (86b93e1)
  • improve transcript readability (48969d4)
  • make new image attachments clickable (81abfc0)
  • normalize EXIF orientation before VLM image inference (a18e8b8), closes #224
  • prevent Markdown responses from being misclassified as raw code (ccab9c3), closes #182
  • restore notarized release builds (0e8939d)
  • set explicit width constraint for standalone cell in transcript tests (d18f1e1)
  • set max tokens to 32k for qwen models (125ef39)

Performance Improvements

  • instrument streaming flush cadence (3c1fd15), closes #226
  • isolate resource monitor updates to sidebar footer (2bfac22)
  • reduce transcript scrolling CPU usage (4f0a3af)