Topic · reinforcement-learning-from-human-feedback
reinforcement-learning-from-human-feedback
Tracked open-source repos tagged reinforcement-learning-from-human-feedback, sorted by stars.
Repos
3
Total stars
12,426
Avg. stars
4,142
Share
0.00%
Related topics
Topics that frequently appear alongside reinforcement-learning-from-human-feedback on the same repo.
Recent risers
Repos created in the last 90 days, tagged reinforcement-learning-from-human-feedback.
No new repos tagged with this topic in the last 90 days.
- #1
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★ 9,969+9Star change over the last 7 days - #2
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
★ 1,613+1Star change over the last 7 days - #3
A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.
★ 844-1Star change over the last 7 days