Skip to main content
buildradar
Sign in
Topic · rl

rl

Tracked open-source repos tagged rl, sorted by stars.

Repos
45
Total stars
111,909
Avg. stars
2,487
Share
0.01%

Topics that frequently appear alongside rl on the same repo.

Recent risers

Repos created in the last 90 days, tagged rl.

No new repos tagged with this topic in the last 90 days.

  • Llama-Chinese@LlamaChinese

    Llama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用

    14,740+1Star change over the last 7 days
  • tianshou@thu-ml

    An elegant PyTorch deep reinforcement learning library.

    10,945+11Star change over the last 7 days
  • dopamine@google

    Dopamine is a research framework for fast prototyping of reinforcement learning algorithms.

    10,901+4Star change over the last 7 days
  • ART@OpenPipe

    Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

    10,679+57Star change over the last 7 days
  • AReaL@areal-project

    The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

    5,703+19Star change over the last 7 days
  • EasyR1@hiyouga

    EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

    5,136+10Star change over the last 7 days
  • hands-on-modern-rl@walkinglabs

    🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.

    4,145+94Star change over the last 7 days
  • agent-sandbox@kubernetes-sigs

    agent-sandbox enables easy management of isolated, stateful, singleton workloads, ideal for use cases like AI agent runtimes and reinforcement learning (RL).

    3,678+92Star change over the last 7 days
  • AlphaZero_Gomoku@junxiaosong

    An implementation of the AlphaZero algorithm for Gomoku (also called Gobang or Five in a Row)

    3,625+2Star change over the last 7 days
  • rl@pytorch

    A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.

    3,540+12Star change over the last 7 days
  • A training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre-trained agents included.

    2,872+3Star change over the last 7 days
  • Papers-in-100-Lines-of-Code@MaximeVandegar

    Implementation of papers in 100 lines of code.

    2,867+8Star change over the last 7 days
  • muzero-general@werner-duvaud

    MuZero

    2,863+4Star change over the last 7 days
  • Awesome-RL-for-LRMs@TsinghuaC3I

    A Survey of Reinforcement Learning for Large Reasoning Models

    2,481+2Star change over the last 7 days
  • all-rl-algorithms@FareedKhan-dev

    Implementation of all RL algorithms in a simpler way

    1,922+7Star change over the last 7 days
  • PRIME@PRIME-RL

    Scalable RL solution for advanced reasoning of language models

    1,872+3Star change over the last 7 days
  • SimpleVLA-RL@PRIME-RL

    [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning

    1,839+9Star change over the last 7 days
  • Latest Advances on System-2 Reasoning

    1,353+1Star change over the last 7 days
  • understand-r1-zero@sail-sg

    Understanding R1-Zero-Like Training: A Critical Perspective

    1,274+1Star change over the last 7 days
  • diy-llm@datawhalechina

    🎓 系统性大语言模型构建课程|🛠️ 覆盖预训练数据工程、Tokenizer、Transformer、MoE、GPU 编程 (CUDA/Triton)、分布式训练、Scaling Laws、推理优化及对齐 (SFT/RLHF/GRPO)|🚀 6 个渐进式作业 + 代码驱动,建立 LLM 全栈认知体系

    1,260+22Star change over the last 7 days
  • TTRL@PRIME-RL

    [NeurIPS 2025] TTRL: Test-Time Reinforcement Learning

    1,121+5Star change over the last 7 days
  • ARPO@RUC-NLPIR

    [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)

    1,111+3Star change over the last 7 days
  • SDPO@lasgroup

    Reinforcement Learning via Self-Distillation (SDPO)

    1,080+8Star change over the last 7 days
  • judgeval@JudgmentLabs

    The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.

    1,058+1Star change over the last 7 days
  • tensorlake@tensorlakeai

    Tensorlake is a serverless runtime for sandboxes and deploying background agentic applications

    995+1Star change over the last 7 days
  • OpenDerisk@derisk-ai

    AI-Native Risk Intelligence Systems, OpenDeRisk——Your application system risk intelligent manager provides 7* 24-hour comprehensive and in-depth protection.

    968+3Star change over the last 7 days
  • mushroom-rl@MushroomRL

    Python library for Reinforcement Learning.

    944+1Star change over the last 7 days
  • AI-Compass@tingaicompass

    “AI-Compass”将为社区指引在 AI 技术海洋中航行的方向,无论你是初学者还是进阶开发者,都能在这里找到通往 AI 各大方向的路径。旨在帮助开发者系统性地了解 AI 的核心概念、主流技术、前沿趋势,并通过实践掌握从理论到落地的全过程。

    926+12Star change over the last 7 days
  • zeroth-bot@zeroth-robotics

    3D-printed open-source humanoid robot platform for sim-to-real and RL

    820+1Star change over the last 7 days
  • mobilegym@Purewhiter

    [EMNLP 2026] MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Evaluation · Scalable Online RL Training

    777+10Star change over the last 7 days
← Back to topics