← The frontier
Technology & AIJul 14, 2026

LLM Judges Can Be Too Generous When There Is No Reference Answer

LLM judges are increasingly being used to evaluate open-ended model responses, often in no-reference settings where a ground-truth answer is unavailable.

LLM judges are increasingly being used to evaluate open-ended model responses, often in no-reference settings where a ground-truth answer is unavailable. However, can they reliably assess in such evaluation setups? We explore this question in this paper through a two stage…

The frontier is open to all. Sign in to learn this from first principles and save it to your knowledge base.