OpenRLHF/OpenRLHF
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
馃殌 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Appears on
Quick read
Latest capture 2026-09-11 10:53
1 path
Agent instructions and tool configuration found in this repository.
Agent instructions
1 observed capture since 2026-09-11. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
OpenClaw-RL: Train any agent simply by talking
Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.
12 weeks, 26 lessons, 52 quizzes, classic Machine Learning for all
Build a modern LLM from scratch. Every line commented. Explained like we are five.
Self-paced bootcamp on Generative AI. Tutorials on ML fundamentals, Ollama, LLMs, RAGs, LangChain, LangGraph, Fine-tuning, DSPy & AI Agents (CrewAI), (Using ChatGPT, gpt-oss, Claude, Qwen, Gemma, Llama, Gemini)