#agent-security
GitHub repositories that have self-applied the topic "agent-security" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #agent-security (top 19 by stars)
🔁 Build reliable recurring AI-agent systems: 968 resources, 22 operational patterns, 22 loop contracts, 8 runtime starters, an interactive atlas, and a structured dataset.
thewaltero/mythos-sentinelAgent permission firewall for MCP tools, shell, files, and x402/Base payments.
MicroMilo/upstream-radarDSH plugin security and dependency monitoring for DeepSeek Harness: exact vulnerable paths, breaking updates, and Agent follow-up.
Actenon/actenon-scanFind where agent-controlled intent reaches consequential actions without an authority check. Python + TypeScript + Go. Zero-dependency SAST for AI agents.
germankovacevic-lab/agent-audit-gateAn audit gate for AI agent outbound messaging: hold the draft, let a senior agent review it (no private-data leak, prompt-injection resistance), then release. Reference implementation.
Krishita17/agent-memory-poisoningAttack taxonomy, toolkit & defenses for persistent-memory poisoning of LLM agents. Author: Krishita Sanjay Choksi.
bharath31/nomineeThe authorization layer for AI agents — allow/deny/ask policy, human approvals, and tamper-evident receipts on every tool call. Zero-dep, framework-neutral, no SaaS.
Vadale/project-guardianAI Guardian Firewall — a local, user-space, agent-agnostic firewall that mediates an autonomous AI agent's actions (files, shell, network, services) with a deterministic policy boundary, a tamper-evident audit log, and a human-in-the-loop approval cockpit. No kernel modules. Apache-2.0.
BrendenKennedy/claude-for-ai-platformsClaude Code scaffold for building AI platforms securely — agent/LLM security, Kubernetes, SRE, observability, identity, and supply chain, grounded in published framework canon (OWASP, NIST, CIS, SLSA). Data-science lanes included.
taixuquant/sentinel-agent-auditMulti-agent collaboration trust measurement & red-team verification (R0, F1, cost, L0-L3). Offline-reproducible Mock baseline for GOAI 2026 Agent Infra.
roleplay-sh/ai-agent-social-engineering-researchCurated research on AI agent social engineering, manipulated delegation, prompt injection, tool-use boundaries, and agent security evaluation.
ch040602/agent-prompt-injection-zooSource-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.
RudrenduPaul/toolgovernRuntime governance for AI agent tool calls: gates shell, filesystem, network, and credential access before execution.
certior/certior-guardA policy hook for Claude Code
sylvesterkaczmarek/worldstate-checkDeterministic postcondition verification for AI agents and autonomous systems using independent observed state.
mobius-style/mobius-browser-guardAn auditable, bounded mediation layer for Claude in Chrome browser actions — content-blind action gating, with its own failure map published.
0xsl1m/shadowshieldUnified open-source security shield for agentic AI systems — defense-in-depth prompt-injection protection (canary tokens, agent-trace alignment audit, tool-call guarding, PII/secret scanning).
guangxiangdebizi/tool-output-spoofing-labBenchmarking schema-valid false tool observations and defense baselines for tool-using LLM agents.
Matik103/sanctum-runtimeOpen-source trust layer for autonomous AI — gate agent, robot, smart home, and industrial actions before they run. Policies, HITL, Ollama/OpenAI, audit. MIT. npm @sanctum-runtime/sdk
RELATED Other topics · full topics ranking →
#claude-code
1,555#ai-agents
1,156#llm
1,066#claude
936#python
802#ai
737#developer-tools
723#mcp
719#codex
517Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology