
DeepEval
DeepEval is an open-source unit testing framework for LLM applications.
Free + paidWebAPICLIOpen sourceKoreanLLM-basedMultimodal
Visit websiteconfident-ai.com
Compare with book-to-skillExplore DeepEval alternativesOverview
DeepEval is an open-source unit testing framework for LLM applications. It offers a Pytest-like experience for developers to test LLM outputs against 14+ metrics including hallucination, toxicity, and bias. It integrates with the Confident AI platform for visualization, regression testing, and real-time monitoring in production environments.
Key features
- 14+ LLM evaluation metrics
- Pytest-based testing
- Real-time monitoring dashboard
- Synthetic data generator
- Regression testing
- CI/CD integration
Pricing
Use cases
- Unit testing LLM apps
- Pre-deployment validation
- Real-time production monitoring
Who it is for
LLM DevelopersQA EngineersAI Product Managers
Integrations
PytestGitHub ActionsLangChainLlamaIndex
Tags
MLOps
How we verified this
Company, pricing, and feature details come from the primary sources below and our latest verification pass. When sources disagree, the official source and the most recent check win.
Last verified 08/30/2026Verified sources: 1
Alternatives
Tools you can use instead

book-to-skill
Converts technical books and document collections into structured, on-demand skills for AI coding agents.
★ 28.8KOpen source
Developer Tools

GPT-6 Astra
OpenAI frontier model for difficult end-to-end reasoning and professional work.
API
Developer Tools

