#tool-use
GitHub Topic「tool-use」がついているAI関連リポジトリの集計。Topicはリポジトリ作者が自己申告するメタタグで、AI関連の文脈で何が「tool-use」と分類されているかを可視化します。
REPOS #tool-use のRepo (TOP 19 / Stars降順)
Python implementation of coding agent core mechanisms, with architecture comparison across Claude Code / Kimi Code / Codex
Nazim22/leadlineEvidence router & policy engine for coding agents — enforces source choice and proof-of-use through Claude Code hooks, including routes to MCP tools. Local-first; no LLM in the hook loop.
delcenjo/llm-sql-agentLLM-powered agent that answers questions over a SQL database using tools
canyang25/AutoSREAn autonomous, LLM-powered SRE agent that takes production incidents from alert to fix/pulls metrics and logs, diagnoses the root cause, runs remediation, and writes the incident report.
111nathanlar/Agent-Worlds adeelmuzaffar31-tech/due-diligence-agentAI agent that researches any company and produces a structured PDF risk report using Claude's tool-use API
kalyvask/ai-safety-osSame agent action, safe or dangerous by context: email team vs investor; delete scratch vs prod config. A safety layer routes each action: auto, confirm, escalate, block. A reading model auto-executes 0% of unsafe actions vs a baseline's 38% (McNemar p<0.001). Plus a runtime loop: an agent earns or loses autonomy with each counterparty over time.
roleplay-sh/ai-agent-social-engineering-researchCurated research on AI agent social engineering, manipulated delegation, prompt injection, tool-use boundaries, and agent security evaluation.
jeffreyhill-ai/jeffreyhill-aiTrusted financial AI, agentic workflows, and senior AI product/platform architecture.
SingularityCoLabs/parallaxA secure, extensible agent-runtime CLI for terminal workflows, with deterministic permissions, human approvals, file tools, shell execution, and durable sessions.
Ciaran11221/opspilotAgentic IT-ops assistant with a live tool-use trace panel - audits account hygiene and ticket SLA risk against synthetic or user-uploaded CSV data, built on the Claude API
wind33441998/modelhubSelf-hosted AI gateway for Claude Code & Codex — route to DeepSeek, Qwen, Kimi, GLM, Gemini, Groq with your own API key. Up to 97% cheaper. 7-day free trial.
xingseq/xingseqA layered Agent Harness framework — ReAct loop, tool registry, multi-provider LLM, workspace isolation. Pure Node.js.
Samuelmartinezduran/agentevalFramework open source para evaluar agentes LLM con function calling / tool use. Scoring por dimensión (tool accuracy, response quality, safety) + CLI, API y dashboard.
ink-charon/codepilot_liteA lightweight local Coding Agent harness inspired by learn-claude-code s20, featuring a minimal LLM tool-use loop, workspace-safe file access, command execution, and pytest-verified Phase 1 tools.
guangxiangdebizi/tool-output-spoofing-labBenchmarking schema-valid false tool observations and defense baselines for tool-using LLM agents.
kalyvask/self-improving-agentic-systemsA controller that learns how a tool-using agent should spend compute at each step (WIDER, DEEPER, DECOMPOSE, STOP, or ESCALATE to a stronger model) and improves its own policy from logged traces. Judged on cost per solved task with paired statistics; ships trained policies and an offline embed API (pip install, no key needed).
the-open-agent/office-tool-useAI tool use for Microsoft Office files
Matik103/sanctum-runtimeOpen-source trust layer for autonomous AI — gate agent, robot, smart home, and industrial actions before they run. Policies, HITL, Ollama/OpenAI, audit. MIT. npm @sanctum-runtime/sdk
RELATED 他のTopicも見る · 全Topicランキング →
#claude-code
1,555#ai-agents
1,156#llm
1,066#claude
936#python
802#ai
737#developer-tools
723#mcp
719#codex
517集計対象: 各Repoの最新contentスナップショットの topics_json に小文字一致でマッチしたもの。 算出方法