AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#ai-security

GitHub repositories that have self-applied the topic "ai-security" — a creator-tagged metadata that surfaces how AI projects describe themselves.

41
tagged repos
162
top 40 stars
10
with tool sigs
40
shown

REPOS Repos for #ai-security (top 40 by stars)

Atrayee-dev/secure-ai-agent-boundary

Secure AI Engineering Framework 2026: Data-Boundary Security for Frontier Models

HTML 115 AI 100
NovaCode37/claude-security-skills

Production-ready Claude Code skills for cybersecurity — secret scanning, SAST, prompt-injection testing, HTTP/JWT/dependency auditing. Zero dependencies.

Python 11 AI 100
Hao610/AI-Model-Atlas

Dual-track AI reference architecture: RAG system design & AI Security Engineering (EN/ZH) | 双轨 AI 参考架构:RAG 系统设计与 AI 安全工程

Python 5 AI 100 Solo live ↗
matterhornso/subscribetome

AI API key & subscription manager for Claude Code — your keys never touch the chat

TypeScript 4 AI 70 Solo live ↗
clay-good/proxilion

Proxilion is the security layer for the agentic workforce. It turns managed AI agents into governed users by enforcing strict cryptographic boundaries on every API call to SaaS like Google Workspace, Salesforce, or Atlassian.

Rust 3 AI 40 Solo live ↗
KezoSec/rag-poisoning-lab

A self-contained AI security lab demonstrating document poisoning, indirect prompt injection, and data exfiltration in RAG systems. Explores the "helpfulness paradox" across local and frontier LLMs.

Python 2 AI 100
Krishita17/agent-memory-poisoning

Attack taxonomy, toolkit & defenses for persistent-memory poisoning of LLM agents. Author: Krishita Sanjay Choksi.

Python 2 AI 100
yushimohuang/agent-watch-approve

🛡️ AI Agent 远程审批系统 - 支持 Cursor/Trae/Claude Code/Codex 等 9 种 AI Agent,通过飞书推送审批卡片到手机/手表/PC,守护你的代码与文件安全

TypeScript 2 AI 100
woshilaohei/border-guard

Border Guard — AI-native territory sovereignty & self-evolving security OS. 6-module D-S fusion engine, trajectory detection, anchor detection, fission engine, territory adjudicator. Full design + Python implementation.

Python 2 AI 70 Solo 1 sig live ↗
sahiee-dev/Sasana

Tamper-evident audit logging for OpenClaw and other local AI agents: hash-chained, locally verifiable, zero cloud dependency.

Python 2 AI 70
Vadale/project-guardian

AI Guardian Firewall — a local, user-space, agent-agnostic firewall that mediates an autonomous AI agent's actions (files, shell, network, services) with a deterministic policy boundary, a tamper-evident audit log, and a human-in-the-loop approval cockpit. No kernel modules. Apache-2.0.

Rust 2 AI 70
Celite12/AI-threat-landscape-2026

A cybersecurity-focused research project analyzing the current state of AI, emerging AI-enabled threats, and practical defensive controls.

1 AI 100
jacobideji/aiiroverlay

AI IR Overlay™ — practical incident response framework for AI agents in production. Built on NIST SP 800-61 r3, mapped to NIST AI RMF, NIST CSF 2.0, OWASP Top 10 for Agentic Applications 2026, ISO/IEC 42001, EU AI Act.

Python 1 AI 100 Solo live ↗
hlsitechio/claude-skills-security

Defensive security audit skills for Claude — tech-stack-keyed and audit-domain-keyed packs for SaaS apps.

Shell 1 AI 100 1 sig
BrendenKennedy/claude-for-ai-platforms

Claude Code scaffold for building AI platforms securely — agent/LLM security, Kubernetes, SRE, observability, identity, and supply chain, grounded in published framework canon (OWASP, NIST, CIS, SLSA). Data-science lanes included.

Shell 1 AI 100 1 sig
Niki-1337/proxy-ai

Open-source AI Security Gateway that sanitizes secrets, PII, and internal context before prompts reach external LLMs.

Rust 1 AI 100
gapilongo/pentest-copilot

Self-hosted pentest copilot. Substrate-first (technique catalog + playbooks + RAG + deterministic tools) with structural verifier rules and LLM-as-judge quality eval. Apache 2.0.

Python 1 AI 100 Solo live ↗
TAIPANBOX/tokenfuse

TokenFuse — runtime control for AI agents: per-run budgets, loop detection, burn forecast, kill-switch. Observability shows the fire; TokenFuse is the automatic extinguisher.

Rust 1 AI 70 1 sig
atuljha-tech/SENTINEL

AI security layer for the agentic internet — sandbox isolation, Civic governance, and a one-API-call clearance system for any OKX.AI agent.

TypeScript 1 AI 70 Solo live ↗
grnbtqdbyx-create/contextforge

Agent context gate for Codex, Claude Code, Copilot, MCP, Cursor, Cline, Gemini and Windsurf repos

TypeScript 1 AI 70 Solo 3 sig live ↗
guorunjie/agentic-workflow-guard

Static analysis for AI automation workflows. Find prompt-injection paths, overpowered tools, and write-capable agent jobs before they run.

JavaScript 1 AI 70 Solo 4 sig live ↗
FelixMa01/agentgate

ai security cli developer-tools agent firewall policy-as-code python anthropic claude-code

Python 1 AI 60
m524security/HIVEBREACH

Autonomous multi-agent AI penetration testing framework — 20 specialist agents, ECC architecture, MITRE ATT&CK + OWASP mapped, HMAC-chained audit trail, Docker sandbox, and multi-LLM backend (Ollama/OpenAI/Anthropic).

Python 1 AI 50 2 sig
roleplay-sh/ai-agent-social-engineering-research

Curated research on AI agent social engineering, manipulated delegation, prompt injection, tool-use boundaries, and agent security evaluation.

0 AI 100
ch040602/agent-prompt-injection-zoo

Source-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.

HTML 0 AI 100 Solo live ↗
rodelgithub/local-llm-bridge

Unlock AI Coding with OpenCode Local Provider 2026 - Auto-Detect Ollama LM Studio

HTML 0 AI 100
MAUROCERON/ai-agent-security-mini-audit

Free AI-agent risk self-check, security checklist, and USD 59 launch-readiness mini-audit offer.

HTML 0 AI 100 Solo live ↗
siddhivinayak-sk/ai-artifact-risk-validator

Validates AI artifacts for security, performance, quality, compliance, and operational risks before peer sharing. It implements a risk framework covering 198 risks across 14 artifact types and 14 scanner modules, including dynamic runtime analysis of live MCP servers.

Python 0 AI 100 Solo 1 sig live ↗
natvernier/ai-trend-radar-lab

A foresight and learning repository for translating early AI, cybersecurity and technical signals into strategic relevance for organisations, governance and adoption.

0 AI 100
rhprasad0/ai-tamperguard

Synthetic Splunk-native dataset and detection prototype for AI-agent monitoring-control-plane tampering

Python 0 AI 100 1 sig
Mithun-veerabuthiran/SSS-Secure-Shielding-Service

Privacy-first Chrome extension and Flask backend for protecting AI prompts through real-time PII detection, anonymization, pseudonymization, and redaction before sensitive data reaches AI services.

0 AI 70 Solo live ↗
Tragentics/heartbeat-control-center

Official desktop companion for tragentics.com — encrypted local vault for agent tokens plus a heartbeat engine that keeps your registered agents online. Windows, macOS, and Linux.

Rust 0 AI 70 live ↗
Bobcatsfan33/loomdb

An agent-native database. Sessions are branches an agent can fork, merge, and rewind; every write records what it was derived from; and taint-and-recall tells you exactly what a poisoned input contaminated. Built on substrate.

Rust 0 AI 70
bylsxy/relayprobe

Evidence-driven GPT-5.5 relay MITM audit console for canaries, prompt injection, tool calls, and usage anomalies.

TypeScript 0 AI 70 Solo 1 sig live ↗
sai-teja-girimaji/dspm-posture-simulator

Live DSPM Posture Simulator — Azure AI agent data exposure risk assessment (educational tool).

HTML 0 AI 70
efanzaluc-prog/monkeycode-sandbox

MonkeyCode 2026: AI Dev Platform with Cloud IDE & Top LLMs - Open Source

HTML 0 AI 70
Builder106/halberd

A JSON-RPC firewall for MCP agents — inspects every tools/call between an LLM and its MCP servers, blocking argument injection and capability creep before they reach the host.

Go 0 AI 70 Solo live ↗
0xsl1m/shadowshield

Unified open-source security shield for agentic AI systems — defense-in-depth prompt-injection protection (canary tokens, agent-trace alignment audit, tool-call guarding, PII/secret scanning).

Python 0 AI 70 Solo live ↗
CrampsP/sentinelforge

Local-first security release checker for authorized code, AI apps, freelancers, and small teams

Python 0 AI 70
rhprasad0/rhprasad0

Ryan Prasad's AI Engineering portfolio and recruiter-agent GitHub profile README

0 AI 60

RELATED Other topics · full topics ranking →

#claude-code

1,564

#ai-agents

1,160

#llm

1,071

#claude

943

#python

806

#ai

741

#developer-tools

728

#mcp

727

#codex

521

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology