Skip to main content
buildradar
Sign in
Topic · reinforcement-learning

reinforcement-learning

Tracked open-source repos tagged reinforcement-learning, sorted by stars.

304 repos
  • awesome-mlss@awesome-mlss

    🤖 Machine Learning Summer School Guide

    3,033+3Star change over the last 7 days
  • agents@tensorflow

    TF-Agents: A reliable, scalable and easy to use TensorFlow library for Contextual Bandits and Reinforcement Learning.

    3,026+0Star change over the last 7 days
  • Deep Learning and deep reinforcement learning research papers and some codes

    3,026+2Star change over the last 7 days
  • mjlab@mujocolab

    Isaac Lab API, powered by MuJoCo-Warp, for RL and robotics research

    2,960+82Star change over the last 7 days
  • openarm@enactic

    A fully open-source humanoid arm for physical AI research and deployment in contact-rich environments.

    2,919+23Star change over the last 7 days
  • A training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre-trained agents included.

    2,873+1Star change over the last 7 days
  • Papers-in-100-Lines-of-Code@MaximeVandegar

    Implementation of papers in 100 lines of code.

    2,870+3Star change over the last 7 days
  • muzero-general@werner-duvaud

    MuZero

    2,865+2Star change over the last 7 days
  • easy-tensorflow@easy-tensorflow

    Simple and comprehensive tutorials in TensorFlow

    2,841+0Star change over the last 7 days
  • YC-Killer@sahibzada-allahyar

    A library of enterprise-grade AI agents designed to democratize artificial intelligence and provide free, open-source alternatives to overvalued Y Combinator startups.

    2,806+7Star change over the last 7 days
  • RAGEN@mll-lab-nu

    Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics

    2,786+6Star change over the last 7 days
  • mctx@google-deepmind

    Monte Carlo tree search in JAX

    2,657+3Star change over the last 7 days
  • PPOxFamily@opendilab

    PPO x Family DRL Tutorial Course(决策智能入门级公开课:8节课帮你盘清算法理论,理顺代码逻辑,玩转决策AI应用实践 )

    2,616+2Star change over the last 7 days
  • Awesome Deep Learning papers for industrial Search, Recommendation and Advertisement. They focus on Embedding, Matching, Pre-Ranking, Ranking, Post Ranking, Relevance, LLM and RL. Please cite our paper "Deep Learning to Rank in Industrial Search Engines, Recommender Systems, and Online Advertising - An Overview and New Perspectives" (TOIS 2026).

    2,589+0Star change over the last 7 days
  • robosuite@ARISE-Initiative

    robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

    2,586+4Star change over the last 7 days
  • Awesome Reasoning LLM Tutorial/Survey/Guide

    2,533+6Star change over the last 7 days
  • Solutions of Reinforcement Learning, An Introduction

    2,433+3Star change over the last 7 days
  • RL4LMs@allenai

    A modular RL library to fine-tune language models to human preferences

    2,395+1Star change over the last 7 days
  • gym-anytrading@AminHP

    The most simple, flexible, and comprehensive OpenAI Gym trading environment (Approved by OpenAI Gym)

    2,387+0Star change over the last 7 days
  • PPO-PyTorch@nikhilbarhate99

    Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch

    2,381+6Star change over the last 7 days
  • ProtoMotions@NVlabs

    ProtoMotions is a GPU-accelerated simulation and learning framework for training physically simulated digital humans and humanoid robots.

    2,359+10Star change over the last 7 days
  • ICCV2019 - Learning to Paint With Model-based Deep Reinforcement Learning

    2,308+2Star change over the last 7 days
  • drl-zh@alessiodm

    Deep Reinforcement Learning: Zero to Hero!

    2,293+0Star change over the last 7 days
  • MimicKit@xbpeng

    A lightweight suite of motion imitation methods for training controllers.

    2,280+15Star change over the last 7 days
  • verl-agent@langfengQ

    verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

    2,274+17Star change over the last 7 days
  • RecSysPapers@tangxyw

    推荐/广告/搜索领域工业界经典以及最前沿论文集合。A collection of industry classics and cutting-edge papers in the field of recommendation/advertising/search.

    2,188-1Star change over the last 7 days
  • gym-pybullet-drones@learnsyslab

    PyBullet Gymnasium environments for single and multi-agent reinforcement learning of quadcopter control

    2,119-2Star change over the last 7 days
  • ASAP@LeCAR-Lab

    [RSS 2025] "ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills"

    2,107+3Star change over the last 7 days
  • diamond@eloialonso

    DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.

    2,100+1Star change over the last 7 days
  • ViZDoom@Farama-Foundation

    Reinforcement Learning environments based on the 1993 game Doom :godmode:

    2,065+0Star change over the last 7 days
← Back to topics