E-03LLM-as-Judge
A model grades outputs against a rubric, standing in for human review at scale.
Further reading
- Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena Zheng et al., NeurIPS 2023
A model grades outputs against a rubric, standing in for human review at scale.