Skip to main content
buildradar
Sign in
Topic · reinforcement-learning

reinforcement-learning

Tracked open-source repos tagged reinforcement-learning, sorted by stars.

304 repos
  • vowpal_wabbit@VowpalWabbit

    Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allreduce, reductions, learning2search, active, and interactive learning.

    8,708+2Star change over the last 7 days
  • pysc2@google-deepmind

    StarCraft II Learning Environment

    8,309+2Star change over the last 7 days
  • PaLM-rlhf-pytorch@lucidrains

    Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

    7,866-1Star change over the last 7 days
  • maths-cs-ai-compendium@HenryNdubuaku

    Become a cracked AI/ML researcher/engineer with this unconventional textbook covering maths, computing, and ML with intuition.

    7,400+21Star change over the last 7 days
  • Reading list for research topics in multimodal machine learning

    6,926+1Star change over the last 7 days
  • A collection of LLM papers, blogs, and projects, with a focus on OpenAI o1 🍓 and reasoning techniques.

    6,902+3Star change over the last 7 days
  • Practical_RL@yandexdataschool

    A course in reinforcement learning in the wild

    6,566+0Star change over the last 7 days
  • 🔬 A curated list of awesome LLMs & deep learning strategies & tools in financial market.

    6,475+23Star change over the last 7 days
  • Mooncake@kvcache-ai

    Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

    6,470+40Star change over the last 7 days
  • A curated list of awesome self-supervised methods

    6,414+1Star change over the last 7 days
  • PufferLib@PufferAI

    Puffing up reinforcement learning

    6,321+8Star change over the last 7 days
  • VLM-R1@om-ai-lab

    Solve Visual Understanding with Reinforced VLMs

    6,018+4Star change over the last 7 days
  • rllm@rllm-org

    Democratizing Reinforcement Learning for LLMs

    5,813+5Star change over the last 7 days
  • AReaL@areal-project

    The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

    5,712+9Star change over the last 7 days
  • open_spiel@google-deepmind

    OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.

    5,452+8Star change over the last 7 days
  • An open source quadruped robot pet framework for developing Boston Dynamics-style four-legged robots that are perfect for STEM, coding & robotics education, IoT robotics applications, AI-enhanced robotics application services, research, and DIY robotics kit development.

    5,246+15Star change over the last 7 days
  • xtuner@InternLM

    A Next-Generation Training Engine Built for Ultra-Large MoE Models

    5,188+3Star change over the last 7 days
  • deep-reinforcement-learning@udacity

    Repo for the Deep Reinforcement Learning Nanodegree program

    5,176+0Star change over the last 7 days
  • EasyR1@hiyouga

    EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

    5,141+5Star change over the last 7 days
  • reasoning-from-scratch@rasbt

    Implement a reasoning LLM in PyTorch from scratch, step by step

    5,138+49Star change over the last 7 days
  • deep-rl-class@huggingface

    This repo contains the Hugging Face Deep Reinforcement Learning Course.

    5,006+10Star change over the last 7 days
  • LLM-RL-Visualized@changyeyu

    🌟100+ 原创 LLM / RL 原理图📚,《大模型算法》作者巨献!💥(100+ LLM/RL Algorithm Maps )

    4,838+12Star change over the last 7 days
  • trlx@CarperAI

    A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)

    4,754-1Star change over the last 7 days
  • RLinf@RLinf

    RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI

    4,705+31Star change over the last 7 days
  • dm_control@google-deepmind

    Google DeepMind's software stack for physics-based simulation and Reinforcement Learning environments, using MuJoCo.

    4,681+6Star change over the last 7 days
  • DouZero@kwai

    [ICML 2021] DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning | 斗地主AI

    4,659+5Star change over the last 7 days
  • alpha-zero-general@suragnair

    A clean implementation based on AlphaZero for any game in any framework + tutorial + Othello/Gobang/TicTacToe/Connect4 and more

    4,512+5Star change over the last 7 days
  • awesome-RLHF@opendilab

    A curated list of reinforcement learning with human feedback resources (continually updated)

    4,424+2Star change over the last 7 days
  • ElegantRL@AI4Finance-Foundation

    Massively Parallel Deep Reinforcement Learning. 🔥

    4,359+2Star change over the last 7 days
  • hands-on-modern-rl@walkinglabs

    🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.

    4,199+54Star change over the last 7 days
← Back to topics