reinforcement-learning
Tracked open-source repos tagged reinforcement-learning, sorted by stars.
- #31
Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allreduce, reductions, learning2search, active, and interactive learning.
★ 8,708+2Star change over the last 7 days - #32★ 8,309+2Star change over the last 7 days
- #33
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
★ 7,866-1Star change over the last 7 days - #34
Become a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computing, and ML with intuition.
★ 7,400+21Star change over the last 7 days - #35
Reading list for research topics in multimodal machine learning
★ 6,926+1Star change over the last 7 days - #36
A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.
★ 6,902+3Star change over the last 7 days - #37
A course in reinforcement learning in the wild
★ 6,566+0Star change over the last 7 days - #38
🔬 A curated list of awesome LLMs & deep learning strategies & tools in financial market.
★ 6,475+23Star change over the last 7 days - #39
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
★ 6,470+40Star change over the last 7 days - #40
A curated list of awesome self-supervised methods
★ 6,414+1Star change over the last 7 days - #41★ 6,321+8Star change over the last 7 days
- #42★ 6,018+4Star change over the last 7 days
- #43★ 5,813+5Star change over the last 7 days
- #44★ 5,712+9Star change over the last 7 days
- #45
OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.
★ 5,452+8Star change over the last 7 days - #46
An open source quadruped robot pet framework for developing Boston Dynamics-style four-legged robots that are perfect for STEM, coding & robotics education, IoT robotics applications, AI-enhanced robotics application services, research, and DIY robotics kit development.
★ 5,246+15Star change over the last 7 days - #47★ 5,188+3Star change over the last 7 days
- #48
Repo for the Deep Reinforcement Learning Nanodegree program
★ 5,176+0Star change over the last 7 days - #49★ 5,141+5Star change over the last 7 days
- #50
Implement a reasoning LLM in PyTorch from scratch, step by step
★ 5,138+49Star change over the last 7 days - #51
This repo contains the Hugging Face Deep Reinforcement Learning Course.
★ 5,006+10Star change over the last 7 days - #52
🌟100+ 原创 LLM / RL 原理图📚,《大模型算法》作者巨献!💥(100+ LLM/RL Algorithm Maps )
★ 4,838+12Star change over the last 7 days - #53
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
★ 4,754-1Star change over the last 7 days - #54★ 4,705+31Star change over the last 7 days
- #55
Google DeepMind's software stack for physics-based simulation and Reinforcement Learning environments, using MuJoCo.
★ 4,681+6Star change over the last 7 days - #56
[ICML 2021] DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning | 斗地主AI
★ 4,659+5Star change over the last 7 days - #57
A clean implementation based on AlphaZero for any game in any framework + tutorial + Othello/Gobang/TicTacToe/Connect4 and more
★ 4,512+5Star change over the last 7 days - #58
A curated list of reinforcement learning with human feedback resources (continually updated)
★ 4,424+2Star change over the last 7 days - #59★ 4,359+2Star change over the last 7 days
- #60
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
★ 4,199+54Star change over the last 7 days