v0.1.3
What's Changed
- docs(review): audit the rollout outcome state space by @Luolc in #106
- docs(review): add the retry policy for a fair eval by @Luolc in #107
- feat: give run_rollout observers and run_unit_test eval_env by @Luolc in #108
- fix(swebench_pro): pin core.autocrlf=false in the eval script by @Luolc in #109
- docs(swebench_pro): record the autocrlf pin as a divergence from the reference by @Luolc in #110
- fix(swebench_pro): also pin core.eol=lf in the eval script by @Luolc in #111
- chore(release): 0.1.3 by @Luolc in #112
Full Changelog: v0.1.2...v0.1.3