reinforcement-learning
Tracked open-source repos tagged reinforcement-learning, sorted by stars.
- #151
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning & ReCall: Learning to Reason with Tool Call for LLMs via Reinforcement Learning
★ 1,432+1Star change over the last 7 days - #152
Concise and beautiful algorithms written in Julia
★ 1,426+0Star change over the last 7 days - #153
A curated list of awesome model based RL resources (continually updated)
★ 1,393+0Star change over the last 7 days - #154★ 1,385+2Star change over the last 7 days
- #155
A programmable and highly maneuverable robotic cat for STEM education and AI-enhanced services.
★ 1,372+1Star change over the last 7 days - #156★ 1,367+1Star change over the last 7 days
- #157
Deep Reinforcement Learning for mobile robot navigation in ROS Gazebo simulator. Using Twin Delayed Deep Deterministic Policy Gradient (TD3) neural network, a robot learns to navigate to a random goal point in a simulated environment while avoiding obstacles.
★ 1,362+2Star change over the last 7 days - #158
Modular Deep Reinforcement Learning framework in PyTorch. Companion library of the book "Foundations of Deep Reinforcement Learning".
★ 1,362+0Star change over the last 7 days - #159
PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT-Opt, PointNet..
★ 1,359+1Star change over the last 7 days - #160
The living ecosystem where AI agents complete tasks through workflow loops, improve through iterative execution, are evaluated by mentor agents or humans in the loop, and turn completed work into reusable work experience and data to improve future agents.
★ 1,334+0Star change over the last 7 days - #161
Minimal Deep Q Learning (DQN & DDQN) implementations in Keras
★ 1,329+1Star change over the last 7 days - #162
A standard format for offline reinforcement learning datasets, with popular reference datasets and related utilities
★ 1,294+3Star change over the last 7 days - #163
Paper reading notes on Deep Learning and Machine Learning
★ 1,270+0Star change over the last 7 days - #164
A simple and well styled PPO implementation. Based on my Medium series: https://medium.com/@eyyu/coding-ppo-from-scratch-with-pytorch-part-1-4-613dfc1b14c8.
★ 1,269+1Star change over the last 7 days - #165
RL research on Android devices.
★ 1,240+0Star change over the last 7 days - #166
Generating sets of formulaic alpha (predictive) stock factors via reinforcement learning.
★ 1,213+8Star change over the last 7 days - #167
Training a humanoid robot for locomotion using Reinforcement Learning
★ 1,210+3Star change over the last 7 days - #168
Lecture notes, tutorial tasks including solutions as well as online videos for the reinforcement learning course hosted by Paderborn University
★ 1,192+1Star change over the last 7 days - #169★ 1,188+0Star change over the last 7 days
- #170
[ECCV 2026] MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE
★ 1,177+0Star change over the last 7 days - #171★ 1,155+7Star change over the last 7 days
- #172
Evaluating and reproducing real-world robot manipulation policies (e.g., RT-1, RT-1-X, Octo) in simulation under common setups (e.g., Google Robot, WidowX+Bridge) (CoRL 2024)
★ 1,153+3Star change over the last 7 days - #173
JMLR: OmniSafe is an infrastructural framework for accelerating SafeRL research.
★ 1,150+1Star change over the last 7 days - #174
Advanced evolutionary computation library built directly on top of PyTorch, created at NNAISENSE.
★ 1,144+1Star change over the last 7 days - #175★ 1,134+0Star change over the last 7 days
- #176
:computer: Learn to make machines learn so that you don't have to struggle to program them; The ultimate list
★ 1,128+0Star change over the last 7 days - #177
A Python-based lightweight robot simulator designed for navigation, control, and learning
★ 1,125+6Star change over the last 7 days - #178
Adversarial skill embeddings for training reusable controllers for physically simulated characters.
★ 1,119+1Star change over the last 7 days - #179★ 1,114+3Star change over the last 7 days
- #180
We introduce the Audio Logical Reasoning (ALR) dataset, consisting of 6,446 text-audio annotated samples specifically designed for complex reasoning tasks. Building on this resource, we propose SoundMind, a rule-based reinforcement learning (RL) algorithm tailored to endow audio language models (ALMs) with deep bimodal reasoning abilities.
★ 1,113+0Star change over the last 7 days