Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
A curated list of awesome LLM frameworks, libraries and software.
GitHub stars and default-branch commits for uhub/awesome-llm.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
Build Conversational AI in minutes ⚡️
Go ahead and axolotl questions
Expose your FastAPI endpoints as Model Context Protocol (MCP) tools, with Auth!
Private chat with local GPT with document, images, video, etc. 100% private, Apache 2.0. Supports oLLaMa, Mixtral, llama.cpp, and more. Demo: https://gpt.h2o.ai/ https://gpt-docs.h2o.ai/
Trae Agent is an LLM-based agent for general purpose software engineering tasks.
Graph-Native Infrastructure for Context and Accountable AI Systems
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
Low-code framework for building custom LLMs, neural networks, and other AI models
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.
Incremental engine for long horizon agents 🌟 Star if you like it!
Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Official inference library for Mistral models
Self-evolving memory OS for LLM & AI Agents: ultra-persistent memory, hybrid-retrieval, and cross-task skill reuse, with 35.24% token savings
🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading
The fullstack MCP framework to develop MCP Apps for ChatGPT / Claude & MCP Servers for AI Agents.
An Open-Source Asynchronous Coding Agent
14MB foundation model for tiny devices; phones, wearables, smart home, and robots.
"AutoAgent: Fully-Automated and Zero-Code LLM Agent Framework"
High-speed Large Language Model Serving for Local Deployment
🐚 Python-powered shell. Full-featured, cross-platform and AI-friendly.
Cybersecurity AI (CAI), the framework for AI Security
SeaTunnel is a multimodal, high-performance, distributed, massive data integration tool.
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
Running large language models on a single GPU for throughput-oriented scenarios.
Enterprise-grade, commercial-friendly agentic workflow platform for building next-generation SuperAgents.
Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
Automate your mobile devices with natural language commands - an LLM agnostic mobile Agent 🤖
The programming language for agents
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
KAG is a logical form-guided reasoning and retrieval framework based on OpenSPG engine and LLMs. It is used to build logical reasoning and factual Q&A solutions for professional domain knowledge bases. It can effectively overcome the shortcomings of the traditional RAG vector similarity calculation model.
the LLM vulnerability scanner
Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
PraisonAI 🦞 — Hire a 24/7 AI Workforce. Stop writing boilerplate and start shipping autonomous self-improving agents that research, plan, code, and execute tasks. Deployed in 5 lines of code with built-in memory, RAG, and support for 100+ LLMs.
An Autonomous LLM Agent for Complex Task Solving
Build effective agents using Model Context Protocol and simple workflow patterns
Accessible large language models via k-bit quantization for PyTorch.
Open Source Deep Research Alternative to Reason and Search on Private Data. Written in Python.
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
Build autonomous AI agents in Python.
Evidently is an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.
slime is an LLM post-training framework for RL Scaling.
A sample app for the Retrieval-Augmented Generation pattern running in Azure, using Azure AI Search for retrieval and Azure OpenAI large language models to power ChatGPT-style and Q&A experiences.
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
Adding guardrails to large language models.
Tools for merging pretrained large language models.
Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.