Faithfulness
DeepEval extracts claims from the answer and checks whether each claim is supported by the retrieval context. A low result is a generator finding: the evidence that reached the model did not justify what it said.
A true statement can still be unfaithful when the retrieved context does not contain it. That is the point: this metric tests grounding, not world knowledge.
inputactual_outputretrieval_context




