v2.7.1
v2.7.1 — Market Readiness
Fixed
researchprofile missingllm_judge_provider— causedValueErroron constructionrun_all.pycomparison table omitted RAGTruth and FreshQA rows- README badge test count stale (1166 → 1837)
Added
benchmarks/BENCHMARK_REPORT.md— consolidated public benchmark reporte2e_eval.pyhybrid-mode CLI flags:--scorer-backend,--llm-judge-provider,--llm-judge-model
Changed
- HF Spaces demo pinned to
>=2.7.1 push_to_hf.shreads version frompyproject.toml
Full Changelog: v2.7.0...v2.7.1