DeepEval

Open-source Python framework for unit-testing LLM apps with pytest-style test cases and ready-made metrics (faithfulness, answer relevancy, hallucination, G-Eval), from Confident AI.

Works with AI agents:llms.txtCLI
Ask your AI about this, with this page as the source:ChatGPT ↗Claude ↗Perplexity ↗

Where it fits

How DeepEval itself is built

19 tools, from its own code, website and Product Hunt page.

Who uses it

No maker's live product on record yet, and 3 open-source projects that declare it in their code.

We haven't found a maker's live product that uses DeepEval yet — only the open-source projects below, which declare it in their code.

Open source: a project that declares DeepEval as a dependency in its public code — verifiable, but not necessarily a live product.

Reliability and open issues

Most wanted on GitHubOpen on 2026-10-04 in confident-ai/deepeval, with activity in the last year — issues and feature requests by 👍.

Alternatives to DeepEval

All alternatives by situation →

Questions makers ask about DeepEval

Is DeepEval free?

Yes — there is a free tier a small product can run on; paid use starts at $200/mo (Confident AI Starter). source ↗

Is DeepEval open source or self-hostable?

Open source, and you can self-host it. source ↗

Can AI coding agents work with DeepEval?

It serves an llms.txt docs index; it has an official CLI (deepeval).

Who uses DeepEval?

No maker's live product we track yet; 3 open-source projects declare it in their code. source ↗