NVIDIA-NeMo/RL
Scalable toolkit for efficient model reinforcement
An implementation of model parallel autoregressive transformers on GPUs, based on the Megatron and DeepSpeed libraries
Quick read
Latest capture 2026-08-15 03:05
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
7 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +21
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Scalable toolkit for efficient model reinforcement
Minimalistic large language model 3D-parallelism training
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Ongoing research training transformer models at scale
A framework for few-shot evaluation of language models.
The hub for EleutherAI's work on interpretability and learning dynamics