AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#agent-security

GitHub repositories that have self-applied the topic "agent-security" — a creator-tagged metadata that surfaces how AI projects describe themselves.

19
tagged repos
73
top 19 stars
5
with tool sigs
19
shown

REPOS Repos for #agent-security (top 19 by stars)

ChaoYue0307/awesome-loop-engineering

🔁 Build reliable recurring AI-agent systems: 968 resources, 22 operational patterns, 22 loop contracts, 8 runtime starters, an interactive atlas, and a structured dataset.

Python 50 AI 70 Solo 1 sig live ↗
thewaltero/mythos-sentinel

Agent permission firewall for MCP tools, shell, files, and x402/Base payments.

JavaScript 6 AI 70
MicroMilo/upstream-radar

DSH plugin security and dependency monitoring for DeepSeek Harness: exact vulnerable paths, breaking updates, and Agent follow-up.

TypeScript 4 AI 70 Solo 1 sig live ↗
Actenon/actenon-scan

Find where agent-controlled intent reaches consequential actions without an authority check. Python + TypeScript + Go. Zero-dependency SAST for AI agents.

Python 4 AI 70
germankovacevic-lab/agent-audit-gate

An audit gate for AI agent outbound messaging: hold the draft, let a senior agent review it (no private-data leak, prompt-injection resistance), then release. Reference implementation.

TypeScript 2 AI 100
Krishita17/agent-memory-poisoning

Attack taxonomy, toolkit & defenses for persistent-memory poisoning of LLM agents. Author: Krishita Sanjay Choksi.

Python 2 AI 100
bharath31/nominee

The authorization layer for AI agents — allow/deny/ask policy, human approvals, and tamper-evident receipts on every tool call. Zero-dep, framework-neutral, no SaaS.

TypeScript 2 AI 70 Solo 2 sig live ↗
Vadale/project-guardian

AI Guardian Firewall — a local, user-space, agent-agnostic firewall that mediates an autonomous AI agent's actions (files, shell, network, services) with a deterministic policy boundary, a tamper-evident audit log, and a human-in-the-loop approval cockpit. No kernel modules. Apache-2.0.

Rust 2 AI 70
BrendenKennedy/claude-for-ai-platforms

Claude Code scaffold for building AI platforms securely — agent/LLM security, Kubernetes, SRE, observability, identity, and supply chain, grounded in published framework canon (OWASP, NIST, CIS, SLSA). Data-science lanes included.

Shell 1 AI 100 1 sig
taixuquant/sentinel-agent-audit

Multi-agent collaboration trust measurement & red-team verification (R0, F1, cost, L0-L3). Offline-reproducible Mock baseline for GOAI 2026 Agent Infra.

Python 0 AI 100
roleplay-sh/ai-agent-social-engineering-research

Curated research on AI agent social engineering, manipulated delegation, prompt injection, tool-use boundaries, and agent security evaluation.

0 AI 100
ch040602/agent-prompt-injection-zoo

Source-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.

HTML 0 AI 100 Solo live ↗
RudrenduPaul/toolgovern

Runtime governance for AI agent tool calls: gates shell, filesystem, network, and credential access before execution.

Python 0 AI 70
certior/certior-guard

A policy hook for Claude Code

JavaScript 0 AI 70
sylvesterkaczmarek/worldstate-check

Deterministic postcondition verification for AI agents and autonomous systems using independent observed state.

Python 0 AI 70 Solo live ↗
mobius-style/mobius-browser-guard

An auditable, bounded mediation layer for Claude in Chrome browser actions — content-blind action gating, with its own failure map published.

Python 0 AI 70
0xsl1m/shadowshield

Unified open-source security shield for agentic AI systems — defense-in-depth prompt-injection protection (canary tokens, agent-trace alignment audit, tool-call guarding, PII/secret scanning).

Python 0 AI 70 Solo live ↗
guangxiangdebizi/tool-output-spoofing-lab

Benchmarking schema-valid false tool observations and defense baselines for tool-using LLM agents.

Python 0 AI 70
Matik103/sanctum-runtime

Open-source trust layer for autonomous AI — gate agent, robot, smart home, and industrial actions before they run. Policies, HITL, Ollama/OpenAI, audit. MIT. npm @sanctum-runtime/sdk

TypeScript 0 AI 50 Solo 1 sig live ↗

RELATED Other topics · full topics ranking →

#claude-code

1,555

#ai-agents

1,156

#llm

1,066

#claude

936

#python

802

#ai

737

#developer-tools

723

#mcp

719

#codex

517

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology