rllm-org/rllm
Democratizing Reinforcement Learning for LLMs
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
Appears on
Quick read
Latest capture 2026-08-24 03:02
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-27. Observed captures are shown by default.
Stars from first capture +3
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Democratizing Reinforcement Learning for LLMs
Awesome resources for in-context learning and prompt engineering: Mastery of the LLMs such as ChatGPT, GPT-3, and FlanT5, with up-to-date and cutting-edge updates.
Open-source pre-training implementation of Google's LaMDA in PyTorch. Adding RLHF similar to ChatGPT.
An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
Benchmarking large language models' complex reasoning ability with chain-of-thought prompting