AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#llmops

GitHub repositories that have self-applied the topic "llmops" — a creator-tagged metadata that surfaces how AI projects describe themselves.

37
tagged repos
163
top 37 stars
9
with tool sigs
37
shown

REPOS Repos for #llmops (top 37 by stars)

GetBusbar/busbar

Point your existing SDK at one URL and reach every LLM vendor — with real failover, not a try/except. One static Rust binary.

Rust 106 AI 70 live ↗
anthony-chaudhary/fak

fak — the Fused Agent Kernel: one Go binary that turns a tool-using agent (Claude Code, Codex, Cursor, any OpenAI/Anthropic/MCP client) into a managed agent: cache-stable model traffic, context compaction + crash resume, nanosecond tool-call policy, local GGUF serving with SSD expert offload.

Go 30 AI 50 Solo 7 sig live ↗
DYAI2025/Plumbline

Plumbline — a self-learning, customer-value-governed agile AI agent team for Claude Code. 87 subagents + skills, TDD defense-in-depth gates, Kaizen retros, a four-body adversarial council, and an empirically benchmarked QA harness. Does it hang true?

Shell 6 AI 70 Solo 1 sig live ↗
salmanzafar949/ctxdiff

git diff for your agent's context window. See exactly what your LLM saw — turn by turn, block by block.

Python 6 AI 70
prat3ik/evalbot

EvalBot — local-first chatbot security & quality evaluation (FastAPI + Next.js). Evaluate chatbot answers against your own docs & guidelines with ML/NLP + AI-judge scoring. Apache-2.0.

Python 3 AI 70
IgorGanapolsky/mac-yolo-safeguards

OS-level safeguards layer for AI agent loops — runaway-kill, timeouts & resource limits so coding agents (Claude Code, Cursor, Codex, Antigravity) in YOLO mode can't burn tokens or freeze your Mac

JavaScript 2 AI 70 Solo 3 sig live ↗
Bobcatsfan33/Pharos

The trust control plane for enterprise AI agents — real-time policy verdicts in under 800ms and litigation-grade evidence of every decision. Pharos decides. Pharos proves.

TypeScript 2 AI 70
DmitryDmitriadi/llm-observability-evaluation

Case study: unified observability + evaluation for automation and conversational AI agents. Versioned error catalog, multi-signal confidence, cost meters.

1 AI 100
ngu-gif/genai-role-playbook

GenAI Career Roadmap 2026 🚀 | AI Job Paths & Skills Guide

HTML 1 AI 100
wiktor-cl/aegis-genai-gateway

Enterprise multi-cloud GenAI agent gateway — AWS Bedrock + Azure AI Foundry integrations, policy-based routing, guardrails, cost control and CI eval gate

Python 1 AI 90 1 sig
wlsdks/reactor

Open-source enterprise AI agent platform for centrally governing agents, tools, memory, RAG, approvals, and durable workflows with FastAPI, LangGraph, LangChain, and LangSmith.

Python 1 AI 70 2 sig
TAIPANBOX/tokenfuse

TokenFuse — runtime control for AI agents: per-run budgets, loop detection, burn forecast, kill-switch. Observability shows the fire; TokenFuse is the automatic extinguisher.

Rust 1 AI 70 1 sig
gregoryhorn/hermes-loop-engineering

Hermes Agent starter kit for safe scheduled, stateful AI-agent loops

Python 1 AI 70 Solo live ↗
pinalmdave/Switchyard

Detect when Claude silently falls back from Fable 5 to Opus 4.8, log it to a local tamper-evident ledger, and keep your work on the frontier model.

Python 1 AI 70 Solo 1 sig live ↗
prakulhiremath/SEMANTIX

A Rust-based PostgreSQL extension that makes relational query optimizers natively aware of LLM token costs, semantic entropy, and latency budgets.

Rust 1 AI 70 Solo live ↗
Hal-Hanami/incident-triage-agent

Read-only first-pass incident triage on the Claude Agent SDK + MCP: classifies an alert, retrieves the matching runbook, and proposes a cited first response — or abstains to a human. Measured over four runs: 100% abstention with 0 missed escalations, ~$0.014 per incident.

Python 0 AI 100
wane528/trace2train

Local CLI to turn failed AI agent traces (wrong tool, bad args, over-refusals) from LangSmith/Langfuse into clean SFT/DPO fine-tuning data

Python 0 AI 100 Solo live ↗
brunovicco/verifiable-ai-governance

Vendor-neutral platform for risk-based, evidence-driven and verifiable AI governance, from intake and conditional approvals to runtime assurance.

Python 0 AI 100 1 sig
dvarahq/dvara-spring-ai-demo

Runnable Spring AI 2.0 demo — govern every LLM call, MCP tool call and agent-to-agent hop through the self-hosted Dvara control plane. Approval gates, PII blocking, hash-chained audit.

Java 0 AI 100 live ↗
runcycles/cycles-spring-ai-starter

Spring AI starter for Cycles — runtime budget and action authority for Spring AI agents

Java 0 AI 100 1 sig live ↗
Yacineutt/AI-AgenticSafe

AI AgenticSafe - stop your AI agent from breaking production at 3 a.m. 8 battle-tested doctrines, 3 real post-mortems, zero dependencies. By Yacine Mahboub, Founder of WEVIA.

0 AI 100 Solo live ↗
prathamesh-git9/llm-gateway

Self-hostable LLM inference gateway: policy routing with fallback chains, circuit breakers, semantic caching, per-tenant cost accounting, and Prometheus metrics.

Python 0 AI 100
Rickvai/managing-four-ai-agents

把一台 16G 的 Mac 跑成 Agent 生产系统:四 Agent Harness 全栈解剖——一个金融背景非程序员的 106 天工程记录

0 AI 100
HadirouTamdamba/enterprise-ai-platform

Enterprise AI Platform — build, deploy, monitor and govern AI applications at scale (RAG, Agents, MLOps, LLMOps, Governance)

Python 0 AI 100
Enterprise-Intelligence-Lab/enterprise-digital-brain

AI-powered Enterprise Knowledge Platform for Intelligent Decision Making, Agentic AI, and Enterprise Intelligence.

Python 0 AI 100
gaurav-bhadane/Enterprise_Agentic_AI_Operations_Platform

Enterprise Agentic AI Platform for autonomous pipeline monitoring, root cause analysis, incident retrieval, data quality validation, and self-healing workflow orchestration using LangGraph, FAISS, Vector Search, and Python.

Python 0 AI 100
WWIIITT/enterprise-financial-intelligence-agent

Enterprise financial intelligence AI agent platform with RAG, SEC EDGAR ingestion, FRED macro analysis, SQL analytics, LangGraph orchestration, security guardrails, evaluation, and observability.

JavaScript 0 AI 100 Solo live ↗
vishipayyallore/generative-ai-engineering

Hands-on Generative AI Engineering repository covering LLMs, Prompt Engineering, RAG, Vector Databases, Fine-Tuning, Multimodal AI, AI Agents, LLMOps, Evaluation, Deployment, and production-ready GenAI applications using OpenAI, Hugging Face, LangChain, LangGraph, and modern AI engineering practices.

0 AI 80
thangldw/ragops

Offline regression tests and explainable release gates for RAG systems and AI agents.

Python 0 AI 70
abhay23-AI/raggate

A thin, CI-gated evaluation gate for RAG & LLM systems — golden set, band-based pass/warn/fail gates, LLM-judge or heuristic scorers. pip install raggate

Python 0 AI 70 Solo live ↗
jainanushk8/swarm-pim

SwarmPIM is a production-grade, containerized PIM system engineered with Next.js 14 and FastAPI. It orchestrates a multi-agent AI swarm utilizing Gemini and Grok with dynamic fallback routing. The architecture features an embedded Qdrant vector cache, high-speed Polars batch ingestion, and analytical DuckDB storage optimized via Apache Arrow.

Python 0 AI 70
Jott2121/sabot

Do your agent pipeline's own checks catch planted faults? Measured on LangGraph, CrewAI and AutoGen: median 16.7%. One prompt-level change takes it to 55.0%. Pre-registered spec, Apache-2.0 harness, every raw trace published.

Python 0 AI 70 Solo live ↗
fengjikui/langgraph-memory-inspector

Local-first DevTools for debugging LangGraph checkpoints and agent memory.

Python 0 AI 70
Victoria824/SpanReplay

OpenTelemetry observability and privacy-aware failure replay for production AI agents.

TypeScript 0 AI 70
Tourinhan/Fund-of-Brian

Agentic AI ops architecture for VC dealflow — Claude + MCP orchestrating CRM, file storage and Notion

HTML 0 AI 70
avgoai/aos-workflow-gate

GitHub Action + CLI for replayable CI/PR/release gate decisions - zero-config Self-Test turning checks, scanners, and AI-agent signals into deterministic, tamper-evident PASS/WARN/BLOCK records. Read-only, zero dependencies, Apache-2.0.

Python 0 AI 60
siva010928/agnos-proxy-oss

Self-hosted, OpenAI-compatible AI gateway that splits your control plane (auth, guardrails, budgets, encrypted key vault, cost & observability) from the translation engine - swap Bifrost / LiteLLM / Portkey / Direct per provider at runtime. Own your keys; contain a compromised engine to one in-flight request.

Python 0 AI 60 Solo live ↗

RELATED Other topics · full topics ranking →

#claude-code

1,564

#ai-agents

1,160

#llm

1,071

#claude

943

#python

806

#ai

741

#developer-tools

728

#mcp

727

#codex

521

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology