v1.70.0
·
173 commits
to master
since this release
New Features
63a79b6- eval: conversation tracking and --preserve-failed flag (commit by @macekond)d599be3- gooddata-eval: optional run_metadata_extra on agentic evaluators (commit by @myhoai)
Bug Fixes
2ebc8b0- eval: fix search_tool correctness always scoring 0% (commit by @Tomkess)f847679- eval: harden _args_match against malformed tool-call JSON (commit by @Tomkess)2299bb7- docs: generate method sub-pages for every class in API reference (commit by @czechian)6c7e272- eval: treat missing/null alert trigger as ALWAYS default (commit by @henrynguyengooddata)16b8722- gooddata-eval: delete metrics created during agentic eval runs (commit by @myhoai)