Llm-as-Judge

10 Aug 2026

LLM-as-Judge

When correctness is subjective, grade agent output with a second, structured LLM call instead of eyeballing every run.