Open highlighted repo slot
Put your repository first
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
GitHub projects from awesome lists
Search names, descriptions, topics, tags, and stacks, then tune results by ecosystem, freshness, health, and cross-list signal.
Open highlighted repo slot
Promote a GitHub repo at the top of Awesome repository list views for 7 days.
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
A library of reinforcement learning components and agents
AI Infra / AI Orchestration / AI Control Plane
OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
🦞 Just talk to your agent — it learns and EVOLVES 🧬.
A standard API for multi-agent reinforcement learning environments, with popular reference environments and related utilities
Accelerated deep learning R&D
Tensorforce: a TensorFlow library for applied reinforcement learning
Skywork-R1V is an advanced multimodal AI model series developed by Skywork AI, specializing in vision-language reasoning.
[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
TF-Agents: A reliable, scalable and easy to use TensorFlow library for Contextual Bandits and Reinforcement Learning.
A training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre-trained agents included.
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models
Implementation of all RL algorithms in a simpler way
Continual learning infra for self-improving agents
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
C++-based high-performance parallel environment execution engine (vectorized env) for general RL environments.
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reinforcement Learning
Evaluate and improve models and agents using environments
Modular Reinforcement Learning (RL) library (implemented in PyTorch, JAX, and NVIDIA Warp) with support for Gymnasium/Gym, NVIDIA Isaac Lab, MuJoCo Playground and other environments
The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
🐝 The First Self-Improving agents with RL / Prompting Optimization
Open-source pre-training implementation of Google's LaMDA in PyTorch. Adding RLHF similar to ChatGPT.
Reinforcement learning library(framework) designed for PyTorch, implements DQN, DDPG, A2C, PPO, SAC, MADDPG, A3C, APEX, IMPALA ...
The Machine Learning project including ML/DL projects, notebooks, cheat codes of ML/DL, useful information on AI/AGI and codes or snippets/scripts/tasks with tips.
PettingZoo and Gymnasium bindings for popular reinforcement learning environments outside of Farama
Safe Bayesian Optimization
Self-tuning multi-agent AI trading system. 8-source signal fusion, Bull/Bear/Judge debate on Claude Opus 4.7, Kelly + ATR position sizing. Python · Kalshi + Polymarket adapters.
Multi-agent simulation library in Python