AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#cost-optimization

GitHub repositories that have self-applied the topic "cost-optimization" — a creator-tagged metadata that surfaces how AI projects describe themselves.

13
tagged repos
127
top 13 stars
3
with tool sigs
13
shown

REPOS Repos for #cost-optimization (top 13 by stars)

PROrunner926/copilot-cache-scout

Multi-Agent Code Review Cost Benchmark 2026: Librarian vs Prompt Cache

HTML 115 AI 100
ingridtoulotte/claude-cost-intelligence

See where LLM money goes, why, and how to cut waste. Local-first FinOps + spend intelligence for Claude Code & the Anthropic API. Pure stdlib, zero deps.

Python 4 AI 90 Solo live ↗
Cost-Caliper/caliper

Precision for your AI spend — see exactly where your Claude Code money goes. Local dashboard + optimization skill (caliper.run)

JavaScript 3 AI 70 2 sig live ↗
CodeWithJuber/forgekit

One config for every AI coding agent — cross-tool config + a cognitive substrate (memory, blast-radius, guardrails) for Claude Code, Codex, Cursor, Gemini, Aider, and more.

JavaScript 2 AI 70 Solo 1 sig live ↗
sfc-gh-jkang/legal-doc-ai-demo

Snowflake Cortex demo: 11-lever cost+quality optimization framework for legal-PDF document AI workloads

Python 1 AI 100 1 sig
Reactance0083/pydantic-ai-multi-llm-cost-optimizer

Flagship free FastAPI starter for LLM routing/cost visibility; $29 Gumroad kit adds deployment, hardening, routing, and verification guides.

Python 1 AI 100 Solo live ↗
Jimmynycu/token-efficiency

A Claude Code plugin that shows exactly where your AI coding session wasted tokens — and how to fix it. Flags waste, never quality. Runs locally.

Python 1 AI 70 Solo live ↗
paopao-13/pecs-multi-agent

PECS: 基于 LangGraph 的四角色多智能体任务求解框架(Planner/Executor/Critic/Synthesizer)。WebShop 真实环境 +25pp (25% vs 0%);GAIA 官方 53 题 26.4% vs ReAct 24.5%(McNemar 不显著)。含 AST 沙箱、50000 token 硬预算、FastAPI 限流/混沌/CI/Prometheus。

Python 0 AI 100
manyu-lnmiit/llm-semantic-cache

Semantic caching layer for LLM & agent tool calls — matches near-duplicate prompts by embedding similarity to cut redundant API spend and latency.

Python 0 AI 100 Solo live ↗
bystray/gonka-mcp-server

MCP server for the GONKA network: run cheap LLM inference through the server (free trial or your own key), get multi-model second opinions, and compare live prices — for any AI agent.

Python 0 AI 70 Solo live ↗
kalyvask/self-improving-agentic-systems

A controller that learns how a tool-using agent should spend compute at each step (WIDER, DEEPER, DECOMPOSE, STOP, or ESCALATE to a stronger model) and improves its own policy from logged traces. Judged on cost per solved task with paired statistics; ships trained policies and an offline embed API (pip install, no key needed).

Python 0 AI 70
iamtural/overkill

Audits real Claude Code usage for effort/model waste — reuses ccusage for raw data, adds the judgment layer it can't provide.

0 AI 70
justinwinter/tiered-dispatch

Confidence-gated model routing for coding agents — run work at the cheapest model that can pass verification. SKILL.md, MIT.

0 AI 35

RELATED Other topics · full topics ranking →

#claude-code

1,555

#ai-agents

1,156

#llm

1,066

#claude

936

#python

802

#ai

737

#developer-tools

723

#mcp

719

#codex

517

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology