gpt-oss
Tracked open-source repos tagged gpt-oss, sorted by stars.
Related topics
Topics that frequently appear alongside gpt-oss on the same repo.
Recent risers
Repos created in the last 90 days, tagged gpt-oss.
- #1
Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp
★ 543
- #1
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
★ 180,007+268Star change over the last 7 days - #2★ 90,817+402Star change over the last 7 days
- #3
SGLang is a high-performance serving framework for large language models and multimodal models.
★ 33,538+791Star change over the last 7 days - #4
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
★ 12,885+39Star change over the last 7 days - #5★ 9,383+3Star change over the last 7 days
- #6
Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code
★ 8,347+9Star change over the last 7 days - #7★ 5,188+3Star change over the last 7 days
- #8
A native macOS app that allows users to chat with a local LLM that can respond with information from files, folders and websites on your Mac without installing any other software. Powered by llama.cpp.
★ 3,312-2Star change over the last 7 days - #9
A powerful Zotero AI and MCP plugin with ChatGPT, Gemini 3.7, Claude Fable 5, Claude Opus 5, DeepSeek V4, Grok, OpenRouter, Kimi k3, GLM 5.3, SiliconFlow, GPT-oss, Gemma 4, Qwen 3.8
★ 2,624+6Star change over the last 7 days - #10
TokenSpeed is a speed-of-light LLM inference engine.
★ 2,075+44Star change over the last 7 days - #11
Local AI app and inference engine for agents. Run open-weight LLMs locally — private, 100% offline on your computer. Join our Discord: https://discord.com/invite/8wGSsvmg4V
★ 1,413+20Star change over the last 7 days - #12
Lightweight Agent Workstation for Codex CLI + Claude Code — with task scheduler, git worktree & remote control
★ 914+8Star change over the last 7 days - #13
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
★ 906+24Star change over the last 7 days - #14
Run MoE models bigger than your RAM. Frontier-size MoE on a 12 GB phone, CPU only, lossless, on stock llama.cpp
★ 543—Star change over the last 7 days