The last place
/coststill priced plan usage: the prompt-caching savings line quoted dollars for subscription providers that bill a flat fee.
Added
npm run export:cataloguewrites the shipped model catalogue to
Codeep-web/src/data/catalogue.jsonstraight fromPROVIDERSand the
context/pricing tables, so codeep.dev renders what the client actually
offers instead of a hand-kept list that had drifted to 4 models against a
catalogue of 72.
Fixed
- Prompt-caching "savings" no longer invents money on a plan.
/cost
priced cached tokens at the provider's pay-per-use input rate even on Z.AI's
Coding Plan, MiniMax, Kimi and Qwen — where caching saves latency, not money,
because nothing is billed per token. The cache read/write counts stay (they
are measured); the dollar figure now covers pay-per-use models only and says
so when a session mixes both, and a plan-only session reads "caching saves
latency, not money" instead. Thebilled at 0.1× / 1.25× input ratenotes are
likewise dropped where no per-token billing applies.