跳到主要內容
buildradar
登入
主題 · reinforcement-learning

reinforcement-learning

標記 reinforcement-learning 主題、收錄中的開源專案,依星數排序。

共 305 個專案
  • gym-mtsim@AminHP

    一個通用、靈活且易於使用的模擬器,並附帶適用於 MetaTrader 5 交易平台的 OpenAI Gym 環境(已獲 OpenAI Gym 認可)

    524+0近 7 天星數變化
  • DI-sheep@opendilab

    羊了個羊 + 深度強化學習

    519+0近 7 天星數變化
  • QeRL@NVlabs

    [ICLR 2026] QeRL 支援在單張 H100 GPU 上對 32B LLM 進行強化學習

    518+0近 7 天星數變化
  • poke-env@hsahovic

    Poke-env:Pokemon Showdown 機器人的 Python 介面。

    510+5近 7 天星數變化
  • open-quadruped@adham-elarabawy

    An open-source 3D-printed quadrupedal robot. Intuitive gait generation through 12-DOF Bezier Curves. Full 6-axis body pose manipulation. Custom 3DOF Leg Inverse Kinematics Model accounting for offsets.

    504+2近 7 天星數變化
← 返回主題列表