Preparing for system design interviews?  Try bugzed.com →

DeepEval

JSON →
library 3.9.6 ·python
verified Jun 28, 2026

DeepEval is an LLM evaluation framework that helps developers evaluate any LLM workflow, from simple prompt chains to complex multi-step agents. It provides a suite of metrics for various evaluation aspects like relevancy, faithfulness, hallucination, and agentic task completion. Currently at version 3.9.6, the library maintains a frequent release cadence, often introducing new metrics, test case types, and developer experience improvements.

total hits 14
actors 5 distinct systems
last hit 17d ago AhrefsBot
GPTBot
3
Sogou
1
ByteDance
1
OAI-SearchBot
1
Humans
4

top countries 🇸🇬 Singapore · 🇺🇸 United States · 🇨🇦 Canada · 🇫🇷 France