
DeepEval
Open-source Python framework for unit-testing LLM apps with pytest-style test cases and ready-made metrics (faithfulness, answer relevancy, hallucination, G-Eval), from Confident AI.
Works with AI agents:llms.txtCLIWhere it fits
How DeepEval itself is built
19 tools, from its own code, website and Product Hunt page.
Who uses it
No maker's live product on record yet, and 3 open-source projects that declare it in their code.
We haven't found a maker's live product that uses DeepEval yet — only the open-source projects below, which declare it in their code.
Open source: a project that declares DeepEval as a dependency in its public code — verifiable, but not necessarily a live product.
Reliability and open issues
- Don't push for confidently.ai platform, or give way to silence it👍 16 · opened Mar 2025 · active Aug 2026
- Remove unconditional, uncontrollable output statements👍 12 · opened May 2024 · active Mar 2026
- ValueError: Evaluation LLM outputted an invalid JSON. Please use a better evaluation model.👍 10 · opened Aug 2024 · active Feb 2026
Alternatives to DeepEval
All alternatives by situation →Questions makers ask about DeepEval
Is DeepEval free?
Yes — there is a free tier a small product can run on; paid use starts at $200/mo (Confident AI Starter). source ↗
Is DeepEval open source or self-hostable?
Open source, and you can self-host it. source ↗
Can AI coding agents work with DeepEval?
It serves an llms.txt docs index; it has an official CLI (deepeval).
Who uses DeepEval?
No maker's live product we track yet; 3 open-source projects declare it in their code. source ↗