Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
An awesome list of Agent Harness engineering resources, including GitHub projects, tools, benchmarks, and practical guides.
GitHub stars and default-branch commits for Picrew/awesome-agent-harness.
332 repos currently saved from this list.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Give your agents the power of the Hugging Face ecosystem
AI Observability & Evaluation
Secure, Fast, and Extensible Sandbox runtime for AI agents.
Instant, Concurrent, Secure & Lightweight Sandbox for AI Agents.
Visual testing tool for MCP servers
🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
AI Agent Engineering Platform built on an Open Source TypeScript AI Agent Framework
The fullstack MCP framework to develop MCP Apps for ChatGPT / Claude & MCP Servers for AI Agents.
An Open-Source Asynchronous Coding Agent
Multi-platform SDK for integrating GitHub Copilot Agent into apps and services
Cybersecurity AI (CAI), the framework for AI Security
Open source MCP Servers for AWS
"Context engineering is the delicate art and science of filling the context window with just the right information for the next step." — Andrej Karpathy. A frontier, first-principles handbook inspired by Karpathy and 3Blue1Brown for moving beyond prompt engineering to the wider discipline of context design, orchestration, and optimization.
🤖 AI Gateway | AI Native API Gateway
Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
Specification and documentation for the Model Context Protocol
A meta-skill that designs domain-specific agent teams, defines specialized agents, and generates the skills they use.
PraisonAI 🦞 — Hire a 24/7 AI Workforce. Stop writing boilerplate and start shipping autonomous self-improving agents that research, plan, code, and execute tasks. Deployed in 5 lines of code with built-in memory, RAG, and support for 100+ LLMs.
Build effective agents using Model Context Protocol and simple workflow patterns
Project management skill system for Agents that uses GitHub Issues and Git worktrees for parallel agent execution.
Manage multiple AI terminal agents like Claude Code, Codex, OpenCode, and Amp.
The sandbox agent framework.
OpenShell is the safe, private runtime for autonomous AI agents.
Flexible and powerful framework for managing multiple AI agents and handling complex conversations
✨ Build AI agents and web apps — with a single binary.
Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built in Swift. Fully offline. Open source.
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
Agentic orchestrator for parallel coding agents — plans tasks, spawns agents, and autonomously handles CI fixes, merge conflicts, and code reviews.
Orchestrate sandboxed coding agents in TypeScript with sandcastle.run()
Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.
🧱 easy fast local-first microVM runtime and library
A community driven registry service for Model Context Protocol (MCP) servers.
This is a simple demonstration of more advanced, agentic patterns built on top of the Realtime API.
Plano is an AI-native proxy server and data plane for agentic apps. Smart LLM routing, observability, agent orchestration, and guardrails so you stay focused on your agents core logic.
TypeScript AI agent orchestration framework with dynamic workflows. Describe the goal, not the graph: a coordinator plans the task DAG at runtime and runs it on any LLM (Claude, ChatGPT, Gemini, DeepSeek, or local models).
可私有部署的多租户知识智能体平台:统一 RAG、知识图谱、多智能体、MCP/Skills、沙盒与权限管理。Self-hosted knowledge agent platform for RAG, knowledge graphs and multi-agent workflows.
Demo of a customer service use case implemented with the OpenAI Agents SDK
Persistent memory system for AI coding agents. Agent-agnostic Go binary with SQLite + FTS5, MCP server, HTTP API, CLI, and TUI.
AIOS: AI Agent Operating System
The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—but scores >74% on SWE-bench verified!
🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓
A simple SWE style browser agent framework that achieves SOTA results on long horizon web tasks.
A model-driven approach to building AI agents in just a few lines of code.
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, and CamelAI
Nexent is a zero-code platform for auto-generating production-grade AI agents using Harness Engineering principles — unified tools, skills, memory, and orchestration with built-in constraints, feedback loops, and control planes.
Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.
Orchestration layer for coding agents (Claude Code, Codex)
An open-source Collaborative Multi-Agent OS for transparent, human-in-the-loop task coordination via Matrix rooms.
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.