github Actively maintained

evidentlyai/evidently

Evidently is ​​an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.

Quick read

Stars
7,778
Forks
893
Open issues
293
Commits
2,795

Activity and growth

Latest capture 2026-08-02 03:14

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+255
Stored snapshots
8

Classification

Metadata

Language
Jupyter Notebook
License
Apache-2.0
Default branch
main
Created
2020-11-25
First commit
2020-11-25
Last pushed
2026-05-02
GitHub updated
2026-08-01
Last synced
2026-08-02 03:14
Stack scanned
2026-08-02 03:14
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

8 observed captures since 2026-05-22. Observed captures are shown by default.

Stars from first capture +255

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

openai/evals

Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.

19,090 stars
Python 7 awesome lists

evalplus/evalplus

Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

1,803 stars
Python 2 awesome lists

hidai25/eval-view

Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.

129 stars
Python 1 awesome list

huggingface/lighteval

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

2,501 stars
Python 3 awesome lists

edublancas/sklearn-evaluation

Machine learning model evaluation made easy: plots, tables, HTML reports, experiment tracking and Jupyter notebook analysis.

3 stars
1 awesome list

expectedparrot/edsl

Design, conduct and analyze results of AI-powered surveys and experiments. Simulate social science and market research with large numbers of AI agents and LLMs.

485 stars
Python 1 awesome list