llmops
Tracked open-source repos tagged llmops, sorted by stars.
Related topics
Topics that frequently appear alongside llmops on the same repo.
Recent risers
Repos created in the last 90 days, tagged llmops.
- #1
Waku Waku! Waku Agent is a local-first AI agent harness you actually own, including loop, memory, eval, all in code built to stay legible as it grows.
★ 1,649 - #2
See what your coding agents did and what it cost. Breaks each task down into work steps — tools used, files changed, tests run, time and tokens spent. Local-first dashboard for Claude Code, Codex, OpenCode, and more. No login, no telemetry.
★ 712 - #3
Hands-on, framework-free Colab notebooks for the AI Engineer / Forward Deployed Engineer (FDE) skill set — model APIs, structured output, tool calling, RAG, evals-as-the-spine, agents (loop from scratch, tool design, guardrails, MCP, Skills), fine-tuning vs LoRA, prompt-injection/security, LLMOps, and customer craft. Runs on the free Groq API.
★ 617
- #1
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, PostgreSQL, real-time data APIs, and more.
★ 58,946+0Star change over the last 7 days - #2
The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails, load balancing, and logging [Bedrock, Azure, OpenAI, Anthropic, OpenAI, VertexAI, vLLM, Nvidia NIM]
★ 57,875+0Star change over the last 7 days - #3
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
★ 34,116+0Star change over the last 7 days - #4
Composio powers 1000+ toolkits, tool search, context management, authentication, and a sandboxed workbench to help you build AI agents that turn intent into action.
★ 30,019+0Star change over the last 7 days - #5
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models and data.
★ 27,783+0Star change over the last 7 days - #6
本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)
★ 24,995+0Star change over the last 7 days - #7
Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.
★ 24,767+0Star change over the last 7 days - #8★ 21,859+0Star change over the last 7 days
- #9
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
★ 21,752+0Star change over the last 7 days - #10
<⚡️> SuperAGI - A dev-first open source autonomous AI agent framework. Enabling developers to build, manage & run useful autonomous agents quickly and reliably.
★ 17,668+0Star change over the last 7 days - #11
Python SDK for Agent AI Observability, Monitoring and Evaluation Framework. Includes features like agent, llm and tools tracing, debugging multi-agentic system, self-hosted dashboard and advanced analytics with timeline and execution graph view
★ 16,159+0Star change over the last 7 days - #12★ 15,594+0Star change over the last 7 days
- #13
A blazing fast AI Gateway with integrated guardrails. Route to 1,600+ LLMs, 50+ AI Guardrails with 1 fast & friendly API.
★ 12,882+0Star change over the last 7 days - #14
Run any open-source LLMs, such as DeepSeek and Llama, as OpenAI compatible API endpoint in the cloud.
★ 12,524+0Star change over the last 7 days - #15
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.
★ 11,922+0Star change over the last 7 days - #16★ 11,298+0Star change over the last 7 days
- #17★ 10,254+0Star change over the last 7 days
- #18
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
★ 8,817+0Star change over the last 7 days - #19★ 8,502+0Star change over the last 7 days
- #20
Evidently is an open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. From tabular data to Gen AI. 100+ metrics.
★ 7,881+0Star change over the last 7 days - #21
Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS.
★ 7,763+0Star change over the last 7 days - #22
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
★ 7,411+0Star change over the last 7 days - #23
Plano is an AI-native proxy server and data plane for agentic apps. Smart LLM routing, observability, agent orchestration, and guardrails so you stay focused on your agents core logic.
★ 7,033+0Star change over the last 7 days - #24
ClearML - Auto-Magical CI/CD to streamline your AI workload. Experiment Management, Data Management, Pipeline, Orchestration, Scheduling & Serving in one MLOps/LLMOps solution
★ 6,853+0Star change over the last 7 days - #25
Ship AI Agents to Google Cloud in minutes, not months. Production-ready templates with built-in CI/CD, evaluation, and observability.
★ 6,550+0Star change over the last 7 days - #26
🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓
★ 6,127+0Star change over the last 7 days - #27
An awesome & curated list of best LLMOps tools for developers
★ 5,920+0Star change over the last 7 days - #28
🐢 Open-Source Evaluation & Testing library for LLM Agents
★ 5,801+0Star change over the last 7 days - #29
Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.
★ 5,711+0Star change over the last 7 days - #30★ 5,574+0Star change over the last 7 days