CarperAI/trlx
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
Train transformer language models with reinforcement learning.
Appears on
Quick read
Latest capture 2026-08-03 03:04
4 paths
Agent instructions and tool configuration found in this repository.
8 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture +547
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
Agentic RL Training at Scale
TensorFlow Reinforcement Learning
Reinforcement Learning in PyTorch
No description.
OpenClaw-RL: Train any agent simply by talking