About the tool
Open-source evaluation framework for LLM apps, agents, and RAG.
DeepEval is an Apache-2.0 Python framework with 50+ research-backed metrics covering hallucination, faithfulness, answer relevancy, toxicity, and agentic task completion. It runs evals unit-test style in CI and supports multi-turn and multimodal evaluation.
- Pricing
- Open source
- Category
- Evals
- Maker
- Not publicly supplied
- Provenance
- Indexed by Attest from public information

✓ Attested
✓ Attested