THUDM/slime
slime is an LLM post-training framework for RL Scaling.
Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
Appears on
Quick read
Latest capture 2026-09-28 10:56
74 paths
Agent instructions and tool configuration found in this repository.
Agent instructions
Agent workspace 2
Claude Code 71
50 more paths detected.
1 observed capture since 2026-09-28. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
slime is an LLM post-training framework for RL Scaling.
SGLang is a high-performance serving framework for large language models and multimodal models.
Ongoing research training transformer models at scale
Low-code framework for building custom LLMs, neural networks, and other AI models
Democratizing Reinforcement Learning for LLMs
Go ahead and axolotl questions