github Actively maintained

confident-ai/deepeval

The LLM Evaluation Framework

Quick read

Stars
18,091
Forks
1,892
Open issues
559
Commits
10,221

Activity and growth

Latest capture 2026-09-04 04:24

Stars · last 7 days
+153 +0.9%
Commits · last 7 days
+59 +0.6%
Stars since tracking
+2443
Stored snapshots
100

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2023-08-10
First commit
2023-08-10
Last pushed
2026-09-03
GitHub updated
2026-09-04
Last synced
2026-09-04 04:24
Stack scanned
2026-09-04 04:24
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

100 observed captures since 2026-05-23. Observed captures are shown by default.

Stars from first capture +2443

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

MigoXLab/dingo

Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool

731 stars
Python 2 awesome lists

AgentOps-AI/agentops

Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, and CamelAI

5,810 stars
Python 5 awesome lists

openlit/openlit

Open source platform for AI Engineering: OpenTelemetry-native LLM Observability, GPU Monitoring, Guardrails, Evaluations, Prompt Management, Vault, Playground. 🚀💻 Integrates with 50+ LLM Providers, VectorDBs, Agent Frameworks and GPUs.

2,666 stars
TypeScript 4 awesome lists

modelscope/evalscope

A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.

3,177 stars
Python 3 awesome lists

comet-ml/opik

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

21,290 stars
Python 10 awesome lists