RadarTopicsBuildersWeeklyReads
Open Source Radar
PKU-Alignment/

safe-rlhf

GitHubWebsite

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

1.6kstars
133forks
18issues
Apache-2.0license
2023since
Star historydaily snapshots by VibeCrowd

Collecting history — the radar snapshots this repo daily. The trend line appears after 3 days of data (1 so far).

Alternatives & relatedmatched by topic overlap
SharePost on XLinkedIn
All trending reposRevenue-verified startups →