You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
{{ message }}
Repository navigation
Default 8k token allowance to allow for reasoning budget.
Allow 8K tokens in default decode samples
Give reasoning-enabled models more room to finish before the decode
completion gate fails. Version the larger allowance as protocol 1.3
without changing prompts, sample counts, prefill, or concurrency.
Keep archived receipts labelled against their original protocol defaults
so the change does not relabel historical benchmark settings.
Signed-off-by: Alex Ellis (OpenFaaS Ltd) <alexellis2@gmail.com>