future-agi/future-agi
Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.
Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.
Appears on
Quick read
Latest capture 2026-08-26 03:03
2 paths
Agent instructions and tool configuration found in this repository.
Agent instructions 2
7 observed captures since 2026-05-23. Observed captures are shown by default.
Stars from first capture +22
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.
🐢 Open-Source Evaluation & Testing library for LLM Agents
A generative AI-powered framework for testing virtual agents.
An MCP server that autonomously evaluates web applications.
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
Templates and workflow for generating PRDs, Tech Designs, and MVP and more using LLMs for AI IDEs