Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
A curated list of awesome LLM frameworks, libraries and software.
GitHub stars and default-branch commits for uhub/awesome-llm.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Automated Penetration Testing Agentic Framework Powered by Large Language Models
Unified framework for building enterprise RAG pipelines with small, specialized models
Nano vLLM
RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works with Claude Code, Codex, OpenClaw, or any LLM agent.
Automate Creation of YouTube Shorts using MoviePy.
structured outputs for llms
Open-source, secure environment with real-world tools for enterprise-grade agents.
Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
One portable memory layer for every AI agent: local-first, Markdown-native, user-owned, and self-evolving across apps, tools, and workflows.
Pocket Flow: Codebase to Tutorial
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
Build Conversational AI in minutes ⚡️
Go ahead and axolotl questions
Access large language models from the command-line
Retrieval and Retrieval-augmented LLMs
Expose your FastAPI endpoints as Model Context Protocol (MCP) tools, with Auth!
Private chat with local GPT with document, images, video, etc. 100% private, Apache 2.0. Supports oLLaMa, Mixtral, llama.cpp, and more. Demo: https://gpt.h2o.ai/ https://gpt-docs.h2o.ai/
Trae Agent is an LLM-based agent for general purpose software engineering tasks.
Graph-Native Infrastructure for Context and Accountable AI Systems
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
Low-code framework for building custom LLMs, neural networks, and other AI models
Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
Pocket Flow: 100-line LLM framework. Let Agents build Agents!
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading
An Open-Source Asynchronous Coding Agent
Self-organizing AI second brain for Obsidian + Claude Code. Drop any source and Claude reads, links, and files it into one connected knowledge graph of plain Markdown you own. AI note-taking, personal knowledge management (PKM), and an open-source Notion alternative. Based on Karpathy's LLM Wiki pattern.
Build, Manage and Deploy AI/ML Systems
"AutoAgent: Fully-Automated and Zero-Code LLM Agent Framework"
🐚 Python-powered shell. Full-featured, cross-platform and AI-friendly.
Cybersecurity AI (CAI), the framework for AI Security
🏔️国立台湾大学、新加坡国立大学、早稻田大学、东京大学,中央研究院(台湾)以及中国重点高校及科研机构,社科、经济、数学、博弈论、哲学、系统工程类学术论文等知识库。
ChatRWKV is like ChatGPT but powered by RWKV (100% RNN) language model, and open source.
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
UFO³: Weaving the Digital Agent Galaxy
Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!
Running large language models on a single GPU for throughput-oriented scenarios.
Automate your mobile devices with natural language commands - an LLM agnostic mobile Agent 🤖
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
KAG is a logical form-guided reasoning and retrieval framework based on OpenSPG engine and LLMs. It is used to build logical reasoning and factual Q&A solutions for professional domain knowledge bases. It can effectively overcome the shortcomings of the traditional RAG vector similarity calculation model.
the LLM vulnerability scanner
Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!