AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#prompt-injection

GitHub repositories that have self-applied the topic "prompt-injection" — a creator-tagged metadata that surfaces how AI projects describe themselves.

43
tagged repos
62
top 40 stars
8
with tool sigs
40
shown

REPOS Repos for #prompt-injection (top 40 by stars)

mlcyclops/lucidagentide

Security-First Agentic IDE Coding Harness

TypeScript 15 AI 45 Solo 2 sig live ↗
NovaCode37/claude-security-skills

Production-ready Claude Code skills for cybersecurity — secret scanning, SAST, prompt-injection testing, HTTP/JWT/dependency auditing. Zero dependencies.

Python 11 AI 100
ralfyishere/agent-zero-trust

Zero-trust repo intake for AI coding agents — scan the instruction environment before Claude Code, Cursor, Codex, or Gemini touches a repo. Ships its own false-negative ledger.

Python 4 AI 100
prat3ik/evalbot

EvalBot — local-first chatbot security & quality evaluation (FastAPI + Next.js). Evaluate chatbot answers against your own docs & guidelines with ML/NLP + AI-judge scoring. Apache-2.0.

Python 3 AI 70
grnbtqdbyx-create/trace-to-skill

Codex Issue Radar and maintainer-readiness tooling for AI coding agents.

TypeScript 3 AI 70 Solo 1 sig live ↗
Krishita17/agent-memory-poisoning

Attack taxonomy, toolkit & defenses for persistent-memory poisoning of LLM agents. Author: Krishita Sanjay Choksi.

Python 2 AI 100
KhushiTripathi762/Sentinel-AI

AI-powered cybersecurity assistant for Prompt Injection and Phishing URL Detection.

JavaScript 2 AI 100 Solo live ↗
KezoSec/rag-poisoning-lab

A self-contained AI security lab demonstrating document poisoning, indirect prompt injection, and data exfiltration in RAG systems. Explores the "helpfulness paradox" across local and frontier LLMs.

Python 2 AI 100
germankovacevic-lab/agent-audit-gate

An audit gate for AI agent outbound messaging: hold the draft, let a senior agent review it (no private-data leak, prompt-injection resistance), then release. Reference implementation.

TypeScript 2 AI 100
calionauta/pi-leakguard

seatbelt for my pi.dev agent: blocks secret-file access, redacts creds from output, blocks egress + git leaks.

TypeScript 2 AI 70
Vadale/project-guardian

AI Guardian Firewall — a local, user-space, agent-agnostic firewall that mediates an autonomous AI agent's actions (files, shell, network, services) with a deterministic policy boundary, a tamper-evident audit log, and a human-in-the-loop approval cockpit. No kernel modules. Apache-2.0.

Rust 2 AI 70
bharath31/nominee

The authorization layer for AI agents — allow/deny/ask policy, human approvals, and tamper-evident receipts on every tool call. Zero-dep, framework-neutral, no SaaS.

TypeScript 2 AI 70 Solo 2 sig live ↗
woshilaohei/border-guard

Border Guard — AI-native territory sovereignty & self-evolving security OS. 6-module D-S fusion engine, trajectory detection, anchor detection, fission engine, territory adjudicator. Full design + Python implementation.

Python 2 AI 70 Solo 1 sig live ↗
Kenny27lokku/prompt-integrity-validator

Lint Your Prompts, Ship Better Agents – Prompt Refiner 2026 Rule Engine

HTML 2 AI 45
Mughal-Baig/local-ai-agent

AgentTrail is a local-first AI agent layer for Ollama/local models that shows its work: semantic search, diff-safe edits, receipts, replay, memory, MCP, reports, and trust controls.

JavaScript 1 AI 100 Solo 1 sig live ↗
mthamil107/ai-bot-shield

Drop-in middleware that detects AI bot traffic. Community-maintained signature database + RFC 9421 Web Bot Auth verification. Node + Python + Go. Apache-2.0.

Python 1 AI 100
BrendenKennedy/claude-for-ai-platforms

Claude Code scaffold for building AI platforms securely — agent/LLM security, Kubernetes, SRE, observability, identity, and supply chain, grounded in published framework canon (OWASP, NIST, CIS, SLSA). Data-science lanes included.

Shell 1 AI 100 1 sig
Edward0l1/skill-flare-discover

Best AI Agent Skill Finder 2026 – Multi-Registry Install & Security Labels

HTML 1 AI 70
mohamedzhioua/proofguard

Kill-tested guard skills for AI coding agents , self-invoking quality gates that catch AI failure modes in security, tests, docs, dependencies, and diffs before the agent says "done." Works with Claude Code, Codex, and Cursor.

JavaScript 1 AI 70 Solo live ↗
TAIPANBOX/tokenfuse

TokenFuse — runtime control for AI agents: per-run budgets, loop detection, burn forecast, kill-switch. Observability shows the fire; TokenFuse is the automatic extinguisher.

Rust 1 AI 70 1 sig
guorunjie/agentic-workflow-guard

Static analysis for AI automation workflows. Find prompt-injection paths, overpowered tools, and write-capable agent jobs before they run.

JavaScript 1 AI 70 Solo 4 sig live ↗
hamodywe/promptfence

Static analysis for AI agents in GitHub Actions — finds where attacker-controlled text reaches an agent's prompt, and what that agent is allowed to do with it.

TypeScript 1 AI 60
roleplay-sh/ai-agent-social-engineering-research

Curated research on AI agent social engineering, manipulated delegation, prompt injection, tool-use boundaries, and agent security evaluation.

0 AI 100
ch040602/agent-prompt-injection-zoo

Source-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.

HTML 0 AI 100 Solo live ↗
perpensum/agent-directed-manipulation

A reproducible definition separating agent-directed manipulation from legitimate machine-readable self-presentation. Two mechanically decidable axes, 13 conformance cases.

HTML 0 AI 100 live ↗
SamsonCyber/llm-injection-field-guide

Dark Promptery: 324 LLM prompt-injection techniques, crosswalked to OWASP/MITRE ATLAS/NIST/CWE.

HTML 0 AI 100 Solo live ↗
Gowrav-M/agent-skillguard

Policy-as-code admission controller for AI agent skills and MCP tools. SkillBOM, lockfiles, and supply-chain baselines.

TypeScript 0 AI 100
MAUROCERON/ai-agent-security-mini-audit

Free AI-agent risk self-check, security checklist, and USD 59 launch-readiness mini-audit offer.

HTML 0 AI 100 Solo live ↗
Vedansh5545/llm-shieldbench

Trustworthy AI evaluation tool for testing chatbot safety, reliability, hallucination behavior, privacy risk, and instruction-following quality.

Python 0 AI 100
marion-official/unicode-cleanup

Claude Code slash commands that flag Unicode used outside string literals — a common prompt-injection / homoglyph vector

Python 0 AI 70
Bobcatsfan33/loomdb

An agent-native database. Sessions are branches an agent can fork, merge, and rewind; every write records what it was derived from; and taint-and-recall tells you exactly what a poisoned input contaminated. Built on substrate.

Rust 0 AI 70
SamsonCyber/agentic-dm-gateway

Hermes-inspired security control plane for LLM agents over private DMs: allowlist, PIN, kill switch, rate limits, injection heuristics, secret redaction, audit log.

Python 0 AI 70
rosscyking1115/redteam-foundry

LLM red-team evaluation harness — prompt injection, refusal, leakage and staleness, with cross-judge validation and attack-corpus audits.

Python 0 AI 70
mobius-style/mobius-browser-guard

An auditable, bounded mediation layer for Claude in Chrome browser actions — content-blind action gating, with its own failure map published.

Python 0 AI 70
Eastern-comptrollership272/crucible-

Route coding tasks through an automated 8-agent pipeline for optimized, secure, and production-ready code output in Claude Code.

0 AI 70
certior/certior-guard

A policy hook for Claude Code

JavaScript 0 AI 70
Builder106/halberd

A JSON-RPC firewall for MCP agents — inspects every tools/call between an LLM and its MCP servers, blocking argument injection and capability creep before they reach the host.

Go 0 AI 70 Solo live ↗
yeodh10/prompt-guard

Prompt Injection Guard - a defense-in-depth (rules + LLM) demo that detects and blocks prompt-injection / jailbreak attempts in LLM user input (Streamlit + Claude).

Python 0 AI 70
0xsl1m/shadowshield

Unified open-source security shield for agentic AI systems — defense-in-depth prompt-injection protection (canary tokens, agent-trace alignment audit, tool-call guarding, PII/secret scanning).

Python 0 AI 70 Solo live ↗
guangxiangdebizi/tool-output-spoofing-lab

Benchmarking schema-valid false tool observations and defense baselines for tool-using LLM agents.

Python 0 AI 70

RELATED Other topics · full topics ranking →

#claude-code

1,564

#ai-agents

1,160

#llm

1,071

#claude

943

#python

806

#ai

741

#developer-tools

728

#mcp

727

#codex

521

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology