#agent-safety
GitHub repositories that have self-applied the topic "agent-safety" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #agent-safety (top 12 by stars)
Control Claude Code in VS Code from Telegram: read sessions, answer questions, approve or deny commands, and keep coding agents moving from your phone.
keynv-labs/keynvSelf-host secrets manager with an AI-safety layer. Aliases instead of values; AI agents never see real credentials. Cloud option coming.
woshilaohei/border-guardBorder Guard — AI-native territory sovereignty & self-evolving security OS. 6-module D-S fusion engine, trajectory detection, anchor detection, fission engine, territory adjudicator. Full design + Python implementation.
Nazim22/leadlineEvidence router & policy engine for coding agents — enforces source choice and proof-of-use through Claude Code hooks, including routes to MCP tools. Local-first; no LLM in the hook loop.
Archerkattri/stepbackGit time-travel for AI coding agents. Checkpoint and rewind Claude Code, Codex, aider, or any command—without touching HEAD, staging, or your branch.
shmindmaster/crewscoreFind the safety rules your AI agent prompt forgot. Offline linter — 23 controls, CLI + Action + browser. No API key.
coreyhiggins/blastradiusHow far does this command reach? Guards AI coding agents against shell commands that leave your machine. Zero dependencies.
CodewithJha/mutinyBehavioral fuzz-testing engine for AI agents — policy violations, tool-call verification, regression tests. Adapter #1: OpenAI Agents SDK.
runcycles/cycles-spring-ai-starterSpring AI starter for Cycles — runtime budget and action authority for Spring AI agents
ShopDevX/breakerboxHard spend caps and a kill-switch for the non-LLM actions your AI coding agent takes. A drop-in Claude Code hook.
rbardyla-boop/claude_powerplantTrust-bounded acceptance harness for Claude coding agents — sanitized workspaces, typed tools, isolated oracle evaluation, evidence receipts.
joeyycli/constitution-lint-actionGitHub Action that lints CLAUDE.md-style agent constitution files for missing operational guardrails — spend limits, injection defense, escalation paths, secrets rules
RELATED Other topics · full topics ranking →
#claude-code
1,564#ai-agents
1,160#llm
1,071#claude
943#python
806#ai
741#developer-tools
728#mcp
727#codex
521Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology