🚀 Wayfinder v0.4.0 — LLM-as-a-Judge Evaluation
This release accompanies AI Engineering Fundamentals – Part 5: LLM-as-a-Judge.
What's Included
- LLM-as-a-Judge evaluation framework
- Configurable evaluation criteria and scoring rubrics
- LLM-based response evaluation with structured outputs
- Judge prompts for evaluating AI responses
- Representative LLM evaluation dataset
- Local evaluation runner
- LangSmith evaluation integration
- Updated project documentation and README
This release represents the code used throughout the article.