THUDM/slime
slime is an LLM post-training framework for RL Scaling.
MrlX: A Multi-Agent Reinforcement Learning Framework
Appears on
Quick read
Latest capture 2026-09-03 03:04
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
7 observed captures since 2026-05-23. Observed captures are shown by default.
Stars from first capture +14
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
slime is an LLM post-training framework for RL Scaling.
Democratizing Reinforcement Learning for LLMs
OpenClaw-RL: Train any agent simply by talking
Reinforcement Learning in PyTorch
Deep Reinforcement Learning for Keras.
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)