Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
GitHub projects from awesome lists
Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
Learn it. Build it. Ship it for others.
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
SGLang is a high-performance serving framework for large language models and multimodal models.
FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
Code for Machine Learning for Trading, 3rd edition — from data sourcing to live execution.
The Unity Machine Learning Agents Toolkit (ML-Agents) is an open-source project that enables games and simulations to serve as environments for training intelligent agents using deep reinforcement learning and imitation learning.
The absolute trainer to light up AI agents.
《李宏毅深度学习教程》(李宏毅老师推荐👍,苹果书🍎),PDF下载地址:https://github.com/datawhalechina/leedl-tutorial/releases
Bullet Physics SDK: real-time collision detection and multi-physics simulation for VR, games, visual effects, robotics, machine learning etc.
PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.
A standard API for single-agent reinforcement learning environments, with popular reference environments and related utilities (formerly Gym)
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allreduce, reductions, learning2search, active, and interactive learning.
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
Deep Learning and Reinforcement Learning Library for Scientists and Engineers
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
A Neural Net Training Interface on TensorFlow, with focus on speed + flexibility
Solve Visual Understanding with Reinforced VLMs
Democratizing Reinforcement Learning for LLMs
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
Deep Reinforcement Learning for Keras.
OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.
A Next-Generation Training Engine Built for Ultra-Large MoE Models
Implement a reasoning LLM in PyTorch from scratch, step by step
This repo contains the Hugging Face Deep Reinforcement Learning Course.
🌟100+ 原创 LLM / RL 原理图📚,《大模型算法》作者巨献!💥(100+ LLM/RL Algorithm Maps )