Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Awesome List
Curated list of the best truly open-source AI projects, models, tools, and infrastructure. Daily updated.
GitHub stars and default-branch commits for alvinreal/awesome-opensource-ai.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
12 weeks, 26 lessons, 52 quizzes, classic Machine Learning for all
🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
A Simple and Universal Swarm Intelligence Engine, Predicting Anything. 简洁通用的群体智能引擎,预测万物
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!
Apache Airflow - A platform to programmatically author, schedule, and monitor workflows
Official inference framework for 1-bit LLMs
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
A powerful MCP toolkit for coding, providing semantic retrieval and editing capabilities - the IDE for your agent
A simple screen parsing tool towards pure vision based GUI agent
No description.
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
Toolkit for linearizing PDFs for LLM datasets/training
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.
🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.
Free and Open Source Machine Translation API. Self-hosted, offline capable and easy to setup.
An orchestration platform for the development, production, and observation of data assets.
The Open Context Layer for Data and AI , OpenMetadata is the open platform for building trusted data context and business semantics for humans, AI assistants, and agents.
Open3D: A Modern Library for 3D Data Processing
The Context Platform for your Data and AI Stack
Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
Cybersecurity AI (CAI), the framework for AI Security
PraisonAI 🦞 — Hire a 24/7 AI Workforce. Stop writing boilerplate and start shipping autonomous self-improving agents that research, plan, code, and execute tasks. Deployed in 5 lines of code with built-in memory, RAG, and support for 100+ LLMs.
Fit interpretable models. Explain blackbox machine learning.
A Low-Code MCP Framework for Building Complex and Innovative RAG Pipelines
Probabilistic time series modeling in Python
Ecommerce Search and Discovery - marqo.ai
Amundsen is a metadata driven application for improving the productivity of data analysts, data scientists and engineers when interacting with data.
Your agent in your terminal, equipped with local tools: writes code, uses the terminal, browses the web. Make your own persistent autonomous agent on top!
Optimizing inference proxy for LLMs
The Python Risk Identification Tool for generative AI (PyRIT) is an open source framework built to empower security professionals and engineers to proactively identify risks in generative AI systems.
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
DeepResearchAgent is a hierarchical multi-agent system designed not only for deep research tasks but also for general-purpose task solving. The framework leverages a top-level planning agent to coordinate multiple specialized lower-level agents, enabling automated task decomposition and efficient execution across diverse and complex domains.
🔅 Shapash: User-friendly Explainability and Interpretability to Develop Reliable and Transparent Machine Learning Models
Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
Skywork-R1V is an advanced multimodal AI model series developed by Skywork AI, specializing in vision-language reasoning.
An open-source visual programming environment for battle-testing prompts to LLMs.
PentestAgent is an AI agent framework for black-box security testing, supporting bug bounty, red-team, and penetration testing workflows.
An AI-powered agentic red team framework that automates offensive security operations, from reconnaissance to exploitation to post-exploitation, with zero human intervention.
TFX is an end-to-end platform for deploying production ML pipelines
Enforce the output format (JSON Schema, Regex etc) of a language model
Responsible AI Toolbox is a suite of tools providing model and data exploration and assessment user interfaces and libraries that enable a better understanding of AI systems. These interfaces and libraries empower developers and stakeholders of AI systems to develop and monitor AI more responsibly, and take better data-driven actions.
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
Large-scale LLM inference engine