github Actively maintained

hidai25/eval-view

Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.

1 awesome list

Quick read

Stars
129
Forks
23
Open issues
3
Commits
809

Activity and growth

Latest capture 2026-08-26 03:03

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+22
Stored snapshots
7

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2025-11-17
First commit
2025-11-17
Last pushed
2026-08-23
GitHub updated
2026-08-23
Last synced
2026-08-26 03:03
Stack scanned
2026-08-26 03:03
Archived
No

AI development signals

2 paths

Agent instructions and tool configuration found in this repository.

Agent instructions 2

View paths

Growth history

Tracked growth

7 observed captures since 2026-05-23. Observed captures are shown by default.

Stars from first capture +22

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

future-agi/future-agi

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

1,544 stars
Python 5 awesome lists

Giskard-AI/giskard-oss

🐢 Open-Source Evaluation & Testing library for LLM Agents

5,728 stars
Python 2 awesome lists

awslabs/agent-evaluation

A generative AI-powered framework for testing virtual agents.

370 stars
Python 1 awesome list

modelscope/evalscope

A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.

3,177 stars
Python 3 awesome lists

KhazP/vibe-coding-prompt-template

Templates and workflow for generating PRDs, Tech Designs, and MVP and more using LLMs for AI IDEs

2,860 stars
TypeScript 2 awesome lists