Skip to main content
buildradar
Sign in

voidful/TextRL

@voidful

Implementation of ChatGPT RLHF (Reinforcement Learning with Human Feedback) on any generation model in huggingface's transformer (blommz-176B/bloom/gpt/bart/T5/MetaICL)

Stars
564
Forks
61
Language
Python
License
MIT
Last push
4 months ago
Pythonlanguage-modelpytorchchatgptreinforcement-learningnlpgpt-3rlhfgpt-2controlled-nlgnlg

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.