#cost-optimization
GitHub Topic「cost-optimization」がついているAI関連リポジトリの集計。Topicはリポジトリ作者が自己申告するメタタグで、AI関連の文脈で何が「cost-optimization」と分類されているかを可視化します。
REPOS #cost-optimization のRepo (TOP 13 / Stars降順)
Multi-Agent Code Review Cost Benchmark 2026: Librarian vs Prompt Cache
ingridtoulotte/claude-cost-intelligenceSee where LLM money goes, why, and how to cut waste. Local-first FinOps + spend intelligence for Claude Code & the Anthropic API. Pure stdlib, zero deps.
Cost-Caliper/caliperPrecision for your AI spend — see exactly where your Claude Code money goes. Local dashboard + optimization skill (caliper.run)
CodeWithJuber/forgekitOne config for every AI coding agent — cross-tool config + a cognitive substrate (memory, blast-radius, guardrails) for Claude Code, Codex, Cursor, Gemini, Aider, and more.
sfc-gh-jkang/legal-doc-ai-demoSnowflake Cortex demo: 11-lever cost+quality optimization framework for legal-PDF document AI workloads
Reactance0083/pydantic-ai-multi-llm-cost-optimizerFlagship free FastAPI starter for LLM routing/cost visibility; $29 Gumroad kit adds deployment, hardening, routing, and verification guides.
Jimmynycu/token-efficiencyA Claude Code plugin that shows exactly where your AI coding session wasted tokens — and how to fix it. Flags waste, never quality. Runs locally.
paopao-13/pecs-multi-agentPECS: 基于 LangGraph 的四角色多智能体任务求解框架(Planner/Executor/Critic/Synthesizer)。WebShop 真实环境 +25pp (25% vs 0%);GAIA 官方 53 题 26.4% vs ReAct 24.5%(McNemar 不显著)。含 AST 沙箱、50000 token 硬预算、FastAPI 限流/混沌/CI/Prometheus。
manyu-lnmiit/llm-semantic-cacheSemantic caching layer for LLM & agent tool calls — matches near-duplicate prompts by embedding similarity to cut redundant API spend and latency.
bystray/gonka-mcp-serverMCP server for the GONKA network: run cheap LLM inference through the server (free trial or your own key), get multi-model second opinions, and compare live prices — for any AI agent.
kalyvask/self-improving-agentic-systemsA controller that learns how a tool-using agent should spend compute at each step (WIDER, DEEPER, DECOMPOSE, STOP, or ESCALATE to a stronger model) and improves its own policy from logged traces. Judged on cost per solved task with paired statistics; ships trained policies and an offline embed API (pip install, no key needed).
iamtural/overkillAudits real Claude Code usage for effort/model waste — reuses ccusage for raw data, adds the judgment layer it can't provide.
justinwinter/tiered-dispatchConfidence-gated model routing for coding agents — run work at the cheapest model that can pass verification. SKILL.md, MIT.
RELATED 他のTopicも見る · 全Topicランキング →
#claude-code
1,564#ai-agents
1,160#llm
1,071#claude
943#python
806#ai
741#developer-tools
728#mcp
727#codex
521集計対象: 各Repoの最新contentスナップショットの topics_json に小文字一致でマッチしたもの。 算出方法