Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
Curated list of the best truly open-source AI projects, models, tools, and infrastructure. Daily updated.
GitHub stars and default-branch commits for alvinreal/awesome-opensource-ai.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
Enhanced ChatGPT Clone: Features Agents, MCP, Skills, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching, message search, Code Interpreter, langchain, DALL-E-3, OpenAPI Actions, Functions, Secure Multi-User Auth, Presets, open-source for self-hosting. Active
LLM API 管理 & 分发系统,支持 OpenAI、Azure、Anthropic Claude、Google Gemini、DeepSeek、字节豆包、ChatGLM、文心一言、讯飞星火、通义千问、360 智脑、腾讯混元等主流模型,统一 API 适配,可用于 key 管理与二次分发。单可执行文件,提供 Docker 镜像,一键部署,开箱即用。LLM API management & key redistribution system, unifying multiple providers under a single API. Single binary, Docker-ready, with an English UI.
Open Source AI Platform - AI Chat with advanced features that works with every LLM
SGLang is a high-performance serving framework for large language models and multimodal models.
Search infrastructure for AI
Temporal service
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.
Gas Town - multi-agent workspace manager
Workflow Engine for Kubernetes
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.
Weaviate is an open-source vector database that stores both objects and vectors, allowing for the combination of vector search with structured filtering with the fault tolerance and scalability of a cloud-native database.
The developer platform for on-demand cloud development environments to create software faster and more securely.
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Instant, Concurrent, Secure & Lightweight Sandbox for AI Agents.
Containers for machine learning
🚀 Next Gen Multi-tenant AI One-Stop Solution. Builtin Admin & Billing System. Enterprise-Grade Unified LLM Gateway Support for 200+ Models And 35+ Providers, Load Balacing w/ Priority-base Routing, Cost Management, Chat Share, Cloud Sync, Credit/Subscription Billing, All File Parsing, Web Search, Built-in Model Cache.
🤖 AI Gateway | AI Native API Gateway
Apache Beam is a unified programming model for Batch and Streaming data processing.
A Datacenter Scale Distributed Inference Serving Framework
The Open Source Feature Store for AI/ML
Dynamic, resilient AI orchestration. Coordinate data, models, and compute as you build AI workflows.
Open-source framework for building AI-powered apps in JavaScript, Go, and Python, built and used in production by Google
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
Local-first session search, analytics, insights, and token use statistics for coding agents, supporting Claude Code, Codex, and more than 20 other agents.
⚡️next-generation personal AI assistant powered by LLM, RAG and agent loops, supporting computer-use, browser-use and coding agent, demo: https://demo.openagentai.org
lakeFS - Data version control for your data lake | Git for data
The CLI and skills that turn any coding assistant into an expert at creating, evaluating, and deploying AI agents on Google Cloud.
An MLOps framework to package, deploy, monitor and manage thousands of production machine learning models
Next Generation Agentic Proxy for AI Agents and MCP servers
Machine Learning Pipelines for Kubeflow
Heterogeneous GPU Sharing on Kubernetes
Kubernetes-native Job Queueing
An Open Standard for lineage metadata collection
Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.
Manages Unified Access to Generative AI Services built on Envoy Gateway
Automated Machine Learning on Kubernetes
MLRun is an open source MLOps platform for quickly building and managing continuous ML applications across their lifecycle. MLRun integrates into your development and CI/CD environment and automates the delivery of production data, ML pipelines, and online applications.
Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.
KAI Scheduler is an open source Kubernetes Native scheduler for AI workloads at large scale
Apache YuniKorn Core