github Actively maintained

PKU-Alignment/safe-rlhf

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

1 awesome list

Quick read

Stars
1,611
Forks
133
Open issues
18
Commits
111

Activity and growth

Latest capture 2026-08-03 03:08

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+7
Stored snapshots
6

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2023-05-15
First commit
2023-05-15
Last pushed
2025-11-24
GitHub updated
2026-07-27
Last synced
2026-08-03 03:08
Stack scanned
2026-08-03 03:08
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +7

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

vwxyzjn/cleanrl

High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

10,309 stars
Python 1 awesome list

jianzhnie/LLamaTuner

Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.

620 stars
Python 1 awesome list

RUCAIBox/R1-Searcher

R1-searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

724 stars
Python 1 awesome list

DLR-RM/stable-baselines3

PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.

13,637 stars
Python 5 awesome lists

OptimalScale/LMFlow

An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.

8,486 stars
Python 2 awesome lists