Skip to main content
buildradar
Sign in
Topic · reinforcement-learning

reinforcement-learning

Tracked open-source repos tagged reinforcement-learning, sorted by stars.

304 repos
  • ReCall@Agent-RL

    ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reinforcement Learning

    1,432+1Star change over the last 7 days
  • Concise and beautiful algorithms written in Julia

    1,426+0Star change over the last 7 days
  • A curated list of awesome model based RL resources (continually updated)

    1,393+0Star change over the last 7 days
  • rl_games@Denys88

    RL implementations

    1,385+2Star change over the last 7 days
  • OpenCat-Old@PetoiCamp

    A programmable and highly maneuverable robotic cat for STEM education and AI-enhanced services.

    1,372+1Star change over the last 7 days
  • smac@oxwhirl

    SMAC: The StarCraft Multi-Agent Challenge

    1,367+1Star change over the last 7 days
  • DRL-robot-navigation@reiniscimurs

    Deep Reinforcement Learning for mobile robot navigation in ROS Gazebo simulator. Using Twin Delayed Deep Deterministic Policy Gradient (TD3) neural network, a robot learns to navigate to a random goal point in a simulated environment while avoiding obstacles.

    1,362+2Star change over the last 7 days
  • SLM-Lab@kengz

    Modular Deep Reinforcement Learning framework in PyTorch. Companion library of the book "Foundations of Deep Reinforcement Learning".

    1,362+0Star change over the last 7 days
  • Popular-RL-Algorithms@quantumiracle

    PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT-Opt, PointNet..

    1,359+1Star change over the last 7 days
  • agent-apprenticeship@ray-r-ren

    The living ecosystem where AI agents complete tasks through workflow loops, improve through iterative execution, are evaluated by mentor agents or humans in the loop, and turn completed work into reusable work experience and data to improve future agents.

    1,334+0Star change over the last 7 days
  • Minimal Deep Q Learning (DQN & DDQN) implementations in Keras

    1,329+1Star change over the last 7 days
  • Minari@Farama-Foundation

    A standard format for offline reinforcement learning datasets, with popular reference datasets and related utilities

    1,294+3Star change over the last 7 days
  • Learning-Deep-Learning@patrick-llgc

    Paper reading notes on Deep Learning and Machine Learning

    1,270+0Star change over the last 7 days
  • PPO-for-Beginners@ericyangyu

    A simple and well styled PPO implementation. Based on my Medium series: https://medium.com/@eyyu/coding-ppo-from-scratch-with-pytorch-part-1-4-613dfc1b14c8.

    1,269+1Star change over the last 7 days
  • android_env@google-deepmind

    RL research on Android devices.

    1,240+0Star change over the last 7 days
  • alphagen@ICT-FinD-Lab

    Generating sets of formulaic alpha (predictive) stock factors via reinforcement learning.

    1,213+8Star change over the last 7 days
  • LearningHumanoidWalking@rohanpsingh

    Training a humanoid robot for locomotion using Reinforcement Learning

    1,210+3Star change over the last 7 days
  • Lecture notes, tutorial tasks including solutions as well as online videos for the reinforcement learning course hosted by Paderborn University

    1,192+1Star change over the last 7 days
  • flow@flow-project

    Computational framework for reinforcement learning in traffic control

    1,188+0Star change over the last 7 days
  • MixGRPO@Tencent-Hunyuan

    [ECCV 2026] MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

    1,177+0Star change over the last 7 days
  • Gym@NVIDIA-NeMo

    Evaluate and improve models and agents using environments

    1,155+7Star change over the last 7 days
  • SimplerEnv@simpler-env

    Evaluating and reproducing real-world robot manipulation policies (e.g., RT-1, RT-1-X, Octo) in simulation under common setups (e.g., Google Robot, WidowX+Bridge) (CoRL 2024)

    1,153+3Star change over the last 7 days
  • omnisafe@PKU-Alignment

    JMLR: OmniSafe is an infrastructural framework for accelerating SafeRL research.

    1,150+1Star change over the last 7 days
  • evotorch@nnaisense

    Advanced evolutionary computation library built directly on top of PyTorch, created at NNAISENSE.

    1,144+1Star change over the last 7 days
  • SMARTS@huawei-noah

    Scalable Multi-Agent RL Training School for Autonomous Driving

    1,134+0Star change over the last 7 days
  • :computer: Learn to make machines learn so that you don't have to struggle to program them; The ultimate list

    1,128+0Star change over the last 7 days
  • ir-sim@hanruihua

    A Python-based lightweight robot simulator designed for navigation, control, and learning

    1,125+6Star change over the last 7 days
  • ASE@nv-tlabs

    Adversarial skill embeddings for training reusable controllers for physically simulated characters.

    1,119+1Star change over the last 7 days
  • ARPO@RUC-NLPIR

    [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)

    1,114+3Star change over the last 7 days
  • SoundMind@xid32

    We introduce the Audio Logical Reasoning (ALR) dataset, consisting of 6,446 text-audio annotated samples specifically designed for complex reasoning tasks. Building on this resource, we propose SoundMind, a rule-based reinforcement learning (RL) algorithm tailored to endow audio language models (ALMs) with deep bimodal reasoning abilities.

    1,113+0Star change over the last 7 days
← Back to topics