dpo
Tracked open-source repos tagged dpo, sorted by stars.
Related topics
Topics that frequently appear alongside dpo on the same repo.
Recent risers
Repos created in the last 90 days, tagged dpo.
No new repos tagged with this topic in the last 90 days.
- #1★ 9,383+3Star change over the last 7 days
- #2
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。
★ 5,771+12Star change over the last 7 days - #3
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
★ 4,929+1,365Star change over the last 7 days - #4
Align Anything: Training All-modality Model with Feedback
★ 4,671+3Star change over the last 7 days - #5
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
★ 4,199+54Star change over the last 7 days - #6
A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).
★ 909-1Star change over the last 7 days - #7
[ICCV 2025] Official code of DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
★ 736+1Star change over the last 7 days - #8
🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.
★ 669-1Star change over the last 7 days - #9
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
★ 621+0Star change over the last 7 days - #10
A curated list of papers on reinforcement learning for video generation
★ 591+1Star change over the last 7 days - #11
tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.
★ 581+0Star change over the last 7 days