github Actively maintained

lucidrains/PaLM-rlhf-pytorch

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

1 awesome list

Quick read

Stars
7,868
Forks
673
Open issues
20
Commits
164

Activity and growth

Latest capture 2026-08-24 03:02

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+3
Stored snapshots
6

Classification

Metadata

Language
Python
License
MIT
Default branch
main
Created
2022-12-09
First commit
2022-12-09
Last pushed
2026-07-27
GitHub updated
2026-08-23
Last synced
2026-08-24 03:02
Stack scanned
2026-08-24 03:02
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

6 observed captures since 2026-05-27. Observed captures are shown by default.

Stars from first capture +3

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

rllm-org/rllm

Democratizing Reinforcement Learning for LLMs

5,755 stars
Python 2 awesome lists

EgoAlpha/prompt-in-context-learning

Awesome resources for in-context learning and prompt engineering: Mastery of the LLMs such as ChatGPT, GPT-3, and FlanT5, with up-to-date and cutting-edge updates.

2,247 stars
Jupyter Notebook 2 awesome lists

conceptofmind/LaMDA-rlhf-pytorch

Open-source pre-training implementation of Google's LaMDA in PyTorch. Adding RLHF similar to ChatGPT.

468 stars
Python 2 awesome lists

OptimalScale/LMFlow

An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.

8,486 stars
Python 2 awesome lists

CarperAI/trlx

A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)

4,752 stars
Python 1 awesome list

FranxYao/chain-of-thought-hub

Benchmarking large language models' complex reasoning ability with chain-of-thought prompting

2,774 stars
Jupyter Notebook 2 awesome lists