DeepEval#

Documentation from DeepEval — the open-source LLM evaluation framework by Confident AI.

A pytest-like framework for unit-testing LLM outputs, with metrics for answer relevancy, faithfulness, hallucination, and RAG quality. Complements RAGAS for evaluation-driven development.