Skip to main content
buildradar
Sign in

lucidrains/PaLM-rlhf-pytorch

@lucidrains

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Stars
7,866
Forks
673
Language
Python
License
MIT
Last push
2 months ago
Pythondeep-learningartificial-intelligenceattention-mechanismsreinforcement-learningtransformershuman-feedback

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.