Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
An awesome list of Agent Harness engineering resources, including GitHub projects, tools, benchmarks, and practical guides.
GitHub stars and default-branch commits for Picrew/awesome-agent-harness.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
The agent that grows with you
Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Cursor, Codex, and Gemini CLI: local deterministic AST parsing, every edge explained, no vector store.
🌐 Make websites accessible for AI agents. Automate tasks online with ease.
Lightweight coding agent that runs in your terminal
Model Context Protocol Servers
DeepSeek Harness: Everything is a Plugin.
An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.
AI coding assistant skill (Claude Code, Codex, OpenCode, Cursor, Gemini CLI, and more). Turn any folder of code, SQL schemas, R scripts, shell scripts, docs, papers, images, or videos into a queryable knowledge graph. App code + database schema + infrastructure in one graph.
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
A coding agent for open models like Kimi K3
A programming framework for agentic AI
The best-benchmarked open-source AI memory system. And it's free.
Framework for orchestrating role-playing, autonomous AI agents. By fostering collaborative intelligence, CrewAI empowers agents to work together seamlessly, tackling complex tasks.
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM
Build, run, and manage agent platforms.
Build resilient agents.
Multi-harness agentic plugin marketplace for Claude Code, Codex CLI, Cursor, OpenCode, GitHub Copilot, and Gemini CLI
Self-evolving Context Database for AI Agents. Unify Agent Memory, Knowledge RAG and Skills.
ChatDev 2.0: Dev All through LLM-powered Multi-Agent Collaboration
Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.
A lightweight, powerful framework for multi-agent workflows
Integrate cutting-edge LLM technology quickly and easily into your apps
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
A powerful MCP toolkit for coding, providing semantic retrieval and editing capabilities - the IDE for your agent
The batteries-included agent harness.
🚀 The fast, Pythonic way to build MCP servers and clients.
Beads - A memory upgrade for your coding agent
Specification and documentation for Agent Skills
A hive mind communication platform
The official Python SDK for Model Context Protocol servers and clients
⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and more
Hindsight: Agent Memory That Learns
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
AI Agent Framework, the Pydantic way
Open source agentic operating system
Build reliable customer-facing AI agents with Parlant: an interaction control harness optimized for controlled, consistent, and predictable LLM interactions.
🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
Supercharge Your LLM Application Evaluations 🚀
Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, security risks, prompt injection, data exfiltration, and supply-chain risks in Claude Code, Codex, and MCP skills before you install them.
Secure, Fast, and Extensible Sandbox runtime for AI agents.
Eigent: The Open Source Cowork Desktop - Local and Free Alternative to Claude Cowork and Codex
Personal memory across agents
Open-source, secure environment with real-world tools for enterprise-grade agents.
A framework for few-shot evaluation of language models.
One portable memory layer for every AI agent: local-first, Markdown-native, user-owned, and self-evolving across apps, tools, and workflows.
A framework for building, orchestrating and deploying AI agents and multi-agent workflows with support for Python and .NET.
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.