Skip to main content
buildradar
Sign in

tatsu-lab/alpaca_farm

@tatsu-lab

A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.

Stars
844
Forks
66
Language
Python
License
Apache-2.0
Last push
2 years ago
Pythondeep-learninglarge-language-modelsnatural-language-processinginstruction-followingreinforcement-learning-from-human-feedback

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.