llm-agents
Tracked open-source repos tagged llm-agents, sorted by stars.
- #31
Autoharness — a self-learning skill layer for Claude Code — distills skills from your real sessions, updates them as you work, and prunes the ones that stop getting used. No daemon, no benchmark.
★ 1,397+67Star change over the last 7 days - #32
Autonomous Agents (LLMs) research papers. Updated Daily.
★ 1,370-2Star change over the last 7 days - #33
80+ production-ready AI agent examples in Python — build autonomous agents, multi-agent systems and agentic AI with uAgents, ASI:One, MCP, A2A, LangChain, CrewAI, Gemini, Claude and OpenAI.
★ 1,146+14Star change over the last 7 days - #34
Agent OS: keep specialist agents in a hub, spin up a temporary orchestrator per task. Local-first, works with any model.
★ 1,103-12Star change over the last 7 days - #35★ 1,095+15Star change over the last 7 days
- #36
Pre-submission AI review stress-test for research papers. A Claude Code skill: review, verdict, revise, verify.
★ 1,089+71Star change over the last 7 days - #37
Agent memory for LLMs: 30 runnable Jupyter notebooks covering conversation buffers, vector stores, knowledge graphs, episodic and semantic memory, MemGPT, Mem0, Letta, Zep, Graphiti, LoCoMo benchmarks, and production patterns.
★ 1,034+95Star change over the last 7 days - #38
Open-source autoresearch powered by autonomous coding agents. Run Claude Code, OpenCode, and Codex with grading, shared knowledge, and multi-agent evolution. Accepted at COLM 2026.
★ 957+21Star change over the last 7 days - #39★ 869+0Star change over the last 7 days
- #40
A curated list of amazingly awesome articles, people, applications, software libraries and projects related to the knowledge management space
★ 869+0Star change over the last 7 days - #41
A codex plugin for running optimization loops inside a codebase. It is useful when you have a measurable target and many possible changes to try: test runtime, build speed, bundle size, model loss, Lighthouse scores, memory use, query latency, or any other metric you can print from a script.
★ 836+1Star change over the last 7 days - #42
More is Different. A multi-agent world engine where AI agents live, talk, compete, ally.
★ 815+3Star change over the last 7 days - #43
🏆 Curated, ranked list of AI agent harnesses (100+) — plus an MCP server, llms.txt & JSON so agents can recommend them too. Rescored weekly.
★ 802+66Star change over the last 7 days - #44
[EMNLP 2026] MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Evaluation · Scalable Online RL Training
★ 786+9Star change over the last 7 days - #45
Easily create LLM tools and agents using plain Bash/JavaScript/Python functions.
★ 775+6Star change over the last 7 days - #46
The official repository of "Position: Agentic Evolution is the Path to Evolving LLMs".
★ 775+13Star change over the last 7 days - #47
An automated AI research-paper writer based off Google's PaperOrchestra paper's implementation through a skills - benchmark + autoraters using any coding agent (Claude Code, Cursor, Antigravity, Cline, Aider). No API keys, no LLM SDKs.
★ 651+3Star change over the last 7 days - #48
An open-source LLM based automatically daily news collecting workflow showcase powered by Agently AI application development framework.
★ 640+1Star change over the last 7 days - #49
AgentLab: An open-source framework for developing, testing, and benchmarking web agents on diverse tasks, designed for scalability and reproducibility.
★ 629+1Star change over the last 7 days - #50
全网最全的、持续更新的、最火爆的 100+ 人格 蒸馏skills 合集| 多agent系统 |名人/导演/天涯大神/古籍/二次元/职场/情感全品类
★ 627+46Star change over the last 7 days - #51★ 621+3Star change over the last 7 days
- #52★ 611+2Star change over the last 7 days
- #53
Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
★ 594—Star change over the last 7 days - #54
A curated list of awesome resources for quantitative investment and trading strategies focusing on artificial intelligence and machine learning applications in finance.
★ 593+12Star change over the last 7 days - #55
EMNLP 2026 Main Conference Paper: MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems (https://arxiv.org/abs/2601.14230)
★ 569+25Star change over the last 7 days - #56
A curated, battle-tested AI-coding toolkit: Claude Code plugins, subagent orchestration, quality gates, and ready-to-copy prompts, extracted from real production use.
★ 543+304Star change over the last 7 days - #57★ 509-1Star change over the last 7 days
- #58
The action firewall for AI agents. Enforce policy and human approval before risky tool calls, shell commands, workflows, and production changes, with auditable evidence.
★ 502—Star change over the last 7 days - #59
🌍 AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resource Paper.
★ 502—Star change over the last 7 days