Skip to main content
buildradar
Sign in
Topic · reinforcement-learning

reinforcement-learning

Tracked open-source repos tagged reinforcement-learning, sorted by stars.

305 repos
  • A collection of recent papers on building autonomous agent. Two topics included: RL-based / LLM-based agents.

    760+3Star change over the last 7 days
  • MO-Gymnasium@Farama-Foundation

    Multi-objective Gymnasium environments for reinforcement learning

    760-1Star change over the last 7 days
  • tmrl@trackmania-rl

    Reinforcement Learning for real-time applications

    737+2Star change over the last 7 days
  • stable-baselines3-contrib@Stable-Baselines-Team

    Contrib package for Stable-Baselines3 - Experimental reinforcement learning (RL) code

    736+5Star change over the last 7 days
  • This is a project about deep reinforcement learning autonomous obstacle avoidance algorithm for UAV.

    730+0Star change over the last 7 days
  • Long-RL@NVlabs

    Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)

    730+1Star change over the last 7 days
  • disco_rl@google-deepmind

    Accompanying code for "Discovering State-of-the-art Reinforcement Algorithms" Nature publication

    729+2Star change over the last 7 days
  • rl-agents@eleurent

    Implementations of Reinforcement Learning and Planning algorithms

    726-1Star change over the last 7 days
  • A curated list of awesome exploration RL resources (continually updated)

    723+0Star change over the last 7 days
  • PyPokerEngine@ishikota

    Poker engine for poker AI development in Python

    723+1Star change over the last 7 days
  • UAV-DDPG@fangvv

    Code for paper "Computation Offloading Optimization for UAV-assisted Mobile Edge Computing: A Deep Deterministic Policy Gradient Approach"

    722+1Star change over the last 7 days
  • neuron_poker@dickreuter

    Texas holdem OpenAi gym poker environment with reinforcement learning based on keras-rl. Includes virtual rendering and montecarlo for equity calculation.

    722+1Star change over the last 7 days
  • DeepRetrieval@pat-jj

    [COLM’25] DeepRetrieval — 🔥 Training Search Agent by RLVR with Retrieval Outcome

    719+2Star change over the last 7 days
  • A curated list of Monte Carlo tree search papers with implementations.

    715+1Star change over the last 7 days
  • Matterport3DSimulator@peteanderson80

    AI Research Platform for Reinforcement Learning from Real Panoramic Images.

    713+1Star change over the last 7 days
  • mpc-reinforcement-learning@FilippoAiraldi

    Reinforcement Learning with Model Predictive Control

    710+1Star change over the last 7 days
  • pymarl2@hijkzzz

    Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)

    708-1Star change over the last 7 days
  • A most Frontend Collection and survey of vision-language model papers, and models GitHub repository. Continuous updates.

    708+3Star change over the last 7 days
  • ns3-gym@tkn-tub

    ns3-gym - The Playground for Reinforcement Learning in Networking Research

    694+2Star change over the last 7 days
  • Flow-Factory@X-GenGroup

    A unified framework for easy fine-tuning in Flow-Matching models

    692+9Star change over the last 7 days
  • irl-imitation@yrlu

    Implementation of Inverse Reinforcement Learning (IRL) algorithms in Python/Tensorflow. Deep MaxEnt, MaxEnt, LPIRL

    681+2Star change over the last 7 days
  • RosettaStone@utilForever

    Hearthstone simulator using C++ with some reinforcement learning

    680+1Star change over the last 7 days
  • AI-Toolbox@Svalorzen

    A C++ framework for MDPs and POMDPs with Python bindings

    671+1Star change over the last 7 days
  • BenchMARL@facebookresearch

    BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL algorithms, tasks, and models while being systematically grounded in its two core tenets: reproducibility and standardization.

    659+2Star change over the last 7 days
  • A PyTorch library for building deep reinforcement learning agents.

    657+0Star change over the last 7 days
  • Complete-Life-Cycle-of-a-Data-Science-Project

    653+0Star change over the last 7 days
  • ReinforcementLearning.jl@JuliaReinforcementLearning

    A reinforcement learning package for Julia

    652-1Star change over the last 7 days
  • Latest Advances on Long Chain-of-Thought Reasoning

    647-1Star change over the last 7 days
  • pgx@sotetsuk

    ♟️ Vectorized RL game environments in JAX

    644+4Star change over the last 7 days
  • Seg-Zero@JIA-Lab-research

    Project Page For "Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement"

    640+0Star change over the last 7 days
← Back to topics