Skip to main content
buildradar
Sign in

sail-sg/oat

@sail-sg

🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.

Stars
669
Forks
63
Language
Python
License
Apache-2.0
Last push
7 months ago
Pythonllmalignmentreasoningrlhfppodistributed-traininggrpodpor1-zerodistributed-rldueling-banditsllm-aligmentllm-explorationonline-alignmentonline-rlthompson-sampling

No related intel yet

This repo has not appeared in any of the sources the radar tracks. The collector runs on a schedule — check back once it covers this repo.