LLM-as-judge
Also known as Model-based evaluation, AI judge. This is the canonical page; those names redirect here.
Pairings in the atlas
- Pairwise preference scoringstandardModel output evaluationllm-training-alignment
Also known as Model-based evaluation, AI judge. This is the canonical page; those names redirect here.