Llm-as-Judge
10 Aug 2026
When correctness is subjective, grade agent output with a second, structured LLM call instead of eyeballing every run.
10 Aug 2026
When correctness is subjective, grade agent output with a second, structured LLM call instead of eyeballing every run.