v1.68.0
·
214 commits
to master
since this release
New Features
22b6da6- sdk: add LLM provider test & list-models enablers (commit by @zdenekmusil-gd)0544cac- add support for IP allowlist policies (commit by @DmitriiNikitinGD)6ab0749- eval: add gooddata-eval model-evaluation CLI (Phase 1 + Phase 2 + Langfuse) (commit by @zdenekmusil-gd)c253bdf- eval: multi-model comparison, gd-eval models command, Langfuse enhancements (commit by @zdenekmusil-gd)d48428a- eval: add dashboard_summary test kind via /summary endpointfe1ef7f- eval: robust summary scoring + multi-case example dataset360f274- eval: add --concurrency flag for parallel item evaluation (commit by @zdenekmusil-gd)d7f9d97- eval: carry summary_input through the Langfuse dataset sourceed3303d- eval: set Langfuse trace version to the model id
Bug Fixes
62bf443- eval: resolve active LLM provider setting by type, not fixed ida43d45b- eval: set provider_type in resolve_and_activate default patha05b708- eval: return SummaryInput from the Langfuse source helpere06dfc9- include user and usergroup (commit by @DmitriiNikitinGD)
Chores
f9639cb- bump ruff to ~=0.15.15 and ty to ~=0.0.40 (commit by @hkad98)c6cdda3- api-client: regenerate against staging (commit by @DmitriiNikitinGD)aa60710- eval: drop example summary dataset (now lives in Langfuse)