Sference Switch v0.1.4 (Beta)
Pre-releaseSference Switch v0.1.4
1M-context Sference models now serve their real context window in Claude
Code instead of compacting at ~180k.
Claude Code resolves its compaction window as min(believed_model_window,
CLAUDE_CODE_AUTO_COMPACT_WINDOW). An id it does not recognize is believed
to hold 200k tokens, and CLAUDE_CODE_MAX_CONTEXT_TOKENS — the documented
override — is ignored for ids beginning with "claude-", which every
derived alias uses. No user setting could raise the window.
Claude Code treats an id containing [1m] as a 1M-token model, so models
at or above 1M context are now published under a [1m] id: GLM-5.3,
GLM-5.3-Flash, GLM-5.2, Kimi-K3 and DeepSeek-V4-Flash. Models below 1M
keep their bare id and are never decorated. Each model lists once; the
bare id of a 1M model stays routable but is no longer advertised, so
existing sessions and configured aliases are unaffected.
Pair a [1m] model with CLAUDE_CODE_AUTO_COMPACT_WINDOW=950000 for a 950k
working window.