Awesome

GitHub projects from awesome lists

Search awesome repositories

Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.

Repos indexed
17,733
Awesome lists tracked
130
Current results
31

Find repositories

Start broad, then narrow by ecosystem, freshness, health, and growth.

Clear 1 refinement
Search mode
Tune results
More filters Topics, generated tags, stack, files, age, archive status, and growth.
Ecosystem
Files

Choose a suggestion or use commas to require multiple files.

Health

Uses known first-commit dates.

Momentum
Filters by observed commit-count growth over the repository's latest 7-day capture window. Repositories without a recent baseline are excluded.
Filters by observed GitHub star growth over the repository's latest 7-day capture window. Repositories without a recent baseline are excluded.
Reset filters
31 repos shown
Highlighted

Open highlighted repo slot

Put your repository first

Promote a GitHub repo at the top of Awesome repository list views for 7 days.

langfuse/langfuse

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23

AI dev
Updated
2026-08-11
Lists
11 list mentions
First commit
2023-05-18
History
50 history points
License
NOASSERTION
Issues
769 open
Forks
3,538
Commits
8,448 commits
Star growth, last 7 days
0 0.0%
Commit velocity, last 7 days
0 0.0%
mlflow/mlflow

The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.

AI dev
Updated
2026-08-11
Lists
4 list mentions
First commit
2018-06-05
History
66 history points
License
Apache-2.0
Issues
2,052 open
Forks
6,140
Commits
12,948 commits
Star growth, last 7 days
0 0.0%
Commit velocity, last 7 days
0 0.0%
promptfoo/promptfoo

Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.

AI dev
Updated
2026-07-24
Lists
7 list mentions
First commit
2023-03-31
History
9 history points
License
MIT
Issues
456 open
Forks
2,119
Commits
9,264 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
comet-ml/opik

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

AI dev
Updated
2026-08-11
Lists
9 list mentions
First commit
2024-09-02
History
70 history points
License
Apache-2.0
Issues
184 open
Forks
1,698
Commits
6,688 commits
Star growth, last 7 days
0 0.0%
Commit velocity, last 7 days
0 0.0%
open-compass/opencompass

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Updated
2026-07-31
Lists
3 list mentions
First commit
2023-07-04
History
6 history points
License
Apache-2.0
Issues
358 open
Forks
821
Commits
1,211 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
Helicone/helicone

🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓

AI dev
Updated
2026-07-25
Lists
6 list mentions
First commit
2022-11-29
History
77 history points
License
Apache-2.0
Issues
165 open
Forks
651
Commits
5,479 commits
Star growth, last 7 days
+19 +0.3%
Commit velocity, last 7 days
0 0.0%
coze-dev/coze-loop

Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.

AI dev
Updated
2026-07-21
Lists
1 list mention
First commit
2025-06-26
History
5 history points
License
Apache-2.0
Issues
66 open
Forks
777
Commits
340 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
Kiln-AI/Kiln

Build, Evaluate, and Optimize AI Systems. Includes evals, RAG, agents, fine-tuning, synthetic data generation, dataset management, MCP, and more.

AI dev
Updated
2026-08-01
Lists
1 list mention
First commit
2024-08-02
History
6 history points
License
NOASSERTION
Issues
79 open
Forks
378
Commits
7,051 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
Marker-Inc-Korea/AutoRAG

AutoRAG: Now your agent can find anything in your computer. It gets smarter if you are using it frequently.

AI dev
Updated
2026-08-01
Lists
2 list mentions
First commit
2024-01-10
History
6 history points
License
NOASSERTION
Issues
117 open
Forks
420
Commits
938 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
EvolvingLMMs-Lab/lmms-eval

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

Updated
2026-08-01
Lists
1 list mention
First commit
2023-12-24
History
6 history points
License
NOASSERTION
Issues
43 open
Forks
633
Commits
1,814 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
open-compass/VLMEvalKit

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

Updated
2026-07-31
Lists
2 list mentions
First commit
2023-11-25
History
6 history points
License
Apache-2.0
Issues
273 open
Forks
739
Commits
1,776 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
langwatch/langwatch

The platform for LLM evaluations and AI agent testing

AI dev
Updated
2026-08-01
Lists
4 list mentions
First commit
2023-09-09
History
8 history points
License
Apache-2.0
Issues
771 open
Forks
340
Commits
5,954 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
embeddings-benchmark/mteb

MTEB: State-of-the-art evaluation of embeddings across languages and modalities

Updated
2026-08-01
Lists
1 list mention
First commit
2022-04-05
History
6 history points
License
Apache-2.0
Issues
309 open
Forks
654
Commits
4,381 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
modelscope/evalscope

A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.

AI dev
Updated
2026-08-03
Lists
2 list mentions
First commit
2023-12-07
History
7 history points
License
Apache-2.0
Issues
31 open
Forks
438
Commits
895 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
lmnr-ai/lmnr

Laminar - open-source observability platform purpose-built for AI agents. YC S24.

AI dev
Updated
2026-08-01
Lists
1 list mention
First commit
2024-08-29
History
6 history points
License
Apache-2.0
Issues
111 open
Forks
221
Commits
1,760 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
ianarawjo/ChainForge

An open-source visual programming environment for battle-testing prompts to LLMs.

Updated
2026-06-10
Lists
1 list mention
First commit
2023-03-26
History
6 history points
License
MIT
Issues
69 open
Forks
256
Commits
417 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
huggingface/lighteval

Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends

Updated
2026-06-29
Lists
2 list mentions
First commit
2024-01-26
History
6 history points
License
MIT
Issues
364 open
Forks
521
Commits
586 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
onestardao/WFGY

WFGY is heading toward WFGY 5.0 Polaris Protocol, a major open-source release for AI reasoning, RAG, agents, and real-world workflows. Includes Problem Map, Global Debug Card, WFGY 4.0, and the CFV Easter Egg.

AI dev
Updated
2026-08-02
Lists
1 list mention
First commit
2025-06-04
History
6 history points
License
NOASSERTION
Issues
11 open
Forks
162
Commits
7,157 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
trpc-group/trpc-agent-go

A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

AI dev
Updated
2026-07-24
Lists
3 list mentions
First commit
2025-05-14
History
4 history points
License
Apache-2.0
Issues
162 open
Forks
270
Commits
1,860 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
TIGER-AI-Lab/ClawBench

Open-source benchmark for browser AI agents on daily tasks.

AI dev
Updated
2026-07-31
Lists
1 list mention
First commit
2026-04-10
History
5 history points
License
Apache-2.0
Issues
38 open
Forks
30
Commits
397 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
INGInious/INGInious

INGInious is a secure and automated exercises assessment platform using your own tests, also providing a pluggable interface with your existing LMS.

Updated
2026-07-30
Lists
1 list mention
First commit
2014-07-01
History
6 history points
License
NOASSERTION
Issues
163 open
Forks
146
Commits
3,219 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history
hidai25/eval-view

Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.

AI dev
Updated
2026-07-03
Lists
1 list mention
First commit
2025-11-17
History
6 history points
License
Apache-2.0
Issues
6 open
Forks
21
Commits
804 commits
Star growth, last 7 days
No 7-day history
Commit velocity, last 7 days
No 7-day history