DeepEval

DeepEval

DeepEval is an open-source unit testing framework for LLM applications.

Free + paidWebAPICLIOpen sourceKoreanLLM-basedMultimodal
Visit websiteconfident-ai.com
Compare with book-to-skillExplore DeepEval alternatives

Overview

DeepEval is an open-source unit testing framework for LLM applications. It offers a Pytest-like experience for developers to test LLM outputs against 14+ metrics including hallucination, toxicity, and bias. It integrates with the Confident AI platform for visualization, regression testing, and real-time monitoring in production environments.

Key features

  • 14+ LLM evaluation metrics
  • Pytest-based testing
  • Real-time monitoring dashboard
  • Synthetic data generator
  • Regression testing
  • CI/CD integration

Pricing

Free + paidStarting price: US$20/mo
View pricing page

Verified on:

Use cases

  • Unit testing LLM apps
  • Pre-deployment validation
  • Real-time production monitoring

Who it is for

LLM DevelopersQA EngineersAI Product Managers

Integrations

PytestGitHub ActionsLangChainLlamaIndex

Tags

MLOps

How we verified this

Company, pricing, and feature details come from the primary sources below and our latest verification pass. When sources disagree, the official source and the most recent check win.

Last verified 08/30/2026Verified sources: 1

Alternatives

Tools you can use instead