reinforcement-learning
Tracked open-source repos tagged reinforcement-learning, sorted by stars.
- #271
This repository collects some codes that encapsulates commonly used algorithms in the field of machine learning. Most of them are based on Numpy, Pandas or Torch. You can deepen your understanding to related model and algorithm or revise it to get the customized code belongs yourself by referring to this repository.
★ 635+0Star change over the last 7 days - #272★ 633+0Star change over the last 7 days
- #273
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
★ 632+9Star change over the last 7 days - #274
API to run VirtualHome, a Multi-Agent Household Simulator
★ 632+2Star change over the last 7 days - #275★ 631+0Star change over the last 7 days
- #276
Explore the Multimodal “Aha Moment” on 2B Model
★ 623+0Star change over the last 7 days - #277
Deploy walk-these-ways project on Unitree Go2
★ 623+3Star change over the last 7 days - #278
Published on Nature Machine Intelligence! The first real robot(quadrotor) based on differentiable physics training.
★ 614+7Star change over the last 7 days - #279
臺灣大學 (NTU) 李宏毅教授「機器學習 (Machine Learning) 2021 Spring 」課程筆記
★ 602+2Star change over the last 7 days - #280
[ICLR 2026] ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
★ 599+4Star change over the last 7 days - #281
Related papers for reinforcement learning, including classic papers and latest papers in top conferences
★ 595+0Star change over the last 7 days - #282★ 593—Star change over the last 7 days
- #283
AAAI 2024 Papers: Explore a comprehensive collection of innovative research papers presented at one of the premier artificial intelligence conferences. Seamlessly integrate code implementations for better understanding. ⭐ experience the forefront of progress in artificial intelligence with this repository!
★ 592+1Star change over the last 7 days - #284
A curated list of papers on reinforcement learning for video generation
★ 591+1Star change over the last 7 days - #285
[ICML 2026] Official resources of "Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning".
★ 591+1Star change over the last 7 days - #286
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
★ 591+10Star change over the last 7 days - #287★ 585+0Star change over the last 7 days
- #288
NeurIPS 2023: Safety-Gymnasium: A Unified Safe Reinforcement Learning Benchmark
★ 580+1Star change over the last 7 days - #289★ 577+1Star change over the last 7 days
- #290
Implementation of ChatGPT RLHF (Reinforcement Learning with Human Feedback) on any generation model in huggingface's transformer (blommz-176B/bloom/gpt/bart/T5/MetaICL)
★ 564+0Star change over the last 7 days - #291
Deep Reinforcement Learning (PPO) in Autonomous Driving (Carla) [from scratch]
★ 563-1Star change over the last 7 days - #292
Meta Agents Research Environments is a comprehensive platform designed to evaluate AI agents in dynamic, realistic scenarios. Unlike static benchmarks, this platform introduces evolving environments where agents must adapt their strategies as new information becomes available, mirroring real-world challenges.
★ 550+3Star change over the last 7 days - #293
A curated list of artificial intelligence resources (Courses, Tools, App, Open Source Project)
★ 548-1Star change over the last 7 days - #294
[NeurIPS 2025 Spotlight] LLM post-training suite — featuring ReasonFlux, ReasonFlux-PRM, and ReasonFlux-Coder.
★ 542+0Star change over the last 7 days - #295★ 537+2Star change over the last 7 days
- #296
UMI on Legs: Making Manipulation Policies Mobile with Manipulation-Centric Whole-body Controllers
★ 537+3Star change over the last 7 days - #297
Curated list of publicly accessible machine learning engineering courses from CalTech, Columbia, Berkeley, MIT, and Stanford.
★ 536+0Star change over the last 7 days - #298
Multi-Objective Reinforcement Learning algorithms implementations.
★ 533+1Star change over the last 7 days - #299
[CVPR 2025 Highlight] InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object Interactions
★ 528+1Star change over the last 7 days - #300
A general-purpose, flexible, and easy-to-use simulator alongside an OpenAI Gym trading environment for MetaTrader 5 trading platform (Approved by OpenAI Gym)
★ 524+0Star change over the last 7 days