FrontierAI.Engineer
Evaluation, Guardrails & Safety

Answer Abstention

Answer abstention evaluates whether a system correctly declines to answer questions it cannot support, rather than fabricating a response. A good benchmark includes unanswerable or out-of-scope questions and rewards the model for saying it does not know. Measuring abstention alongside accuracy captures a safety property that pure correctness metrics miss, since confidently wrong answers can be more harmful than none.