Skip to content

ctx v0.22.0

Choose a tag to compare

@github-actions github-actions released this 06 Apr 22:35
· 1599 commits to root since this release
v0.22.0
d0672c9

v0.22.0 — Dual-Model Dream Architecture

Dream Mode can now use a separate, larger model for background
cross-referencing while synthesis stays on the fast model. Evaluated
10 Gemma models — qwen3.5 family remains unbeaten. qwen3.5:27b is the
only model to pass all 4 dream tests including causal relationship
detection.

Per-function LLM configuration:

  • OLLAMA_DREAM_MODEL (separate model for dream evaluation)
  • OLLAMA_DREAM_NUM_CTX / OLLAMA_CHAT_NUM_CTX (per-function context)
  • OLLAMA_DREAM_THINK / OLLAMA_THINK (per-function think mode)
  • Removed global ThinkMode — think *bool threaded through full chain

Continuous dream loop:

  • Replaced 10s ticker with back-to-back processing
  • 120s idle wait when no blocks available
  • 2s yield to active queries (demand interruption)
  • 10s pause on error to prevent tight loops

Gemma evaluation results (Session 21):

  • 10 models tested (gemma3:4b/12b/27b, gemma3n:e2b/e4b,
    gemma4:e2b/e4b/26b/31b, embeddinggemma)
  • Best Gemma for dream: gemma4:e2b (3/4), but only 80% synthesis KW
  • Larger models consistently worse (gemma3:12b 0/4, gemma4:31b 0/4)
  • gemma4:26b initially scored 0/4 due to markdown fence wrapping
    (format:"json" not respected); actual quality 2/4 after strip

VRAM budget (24 GB Quadro RTX 6000):

  • Query path: qwen3-embedding:8b (5.3 GB) + qwen3.5:9b (9.3 GB) = 14.6 GB
  • Dream path: qwen3.5:27b (22.2 GB at num_ctx=16384) — Ollama swaps

Installation

With Go:

go install github.com/GottZ/ctx/cmd/ctx@v0.22.0

Binary download:
Download the binary for your platform, make it executable, move to PATH:

chmod +x ctx-*
sudo mv ctx-* /usr/local/bin/ctx