#verification
GitHub repositories that have self-applied the topic "verification" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #verification (top 21 by stars)
"Done" means nothing — a menhera-style completion gate for Claude Code that blocks empty completion claims until tests actually ran green and no TODOs are left behind.
pinetreeB/show-me-the-workAI 코딩 비서(Claude·Codex·Gemini)가 확인도 안 하고 '다 됐어요' 하는 걸 자동으로 막아주는 품질 안전장치 — 비개발자·바이브코더를 위한 한국어 우선 하네스.
winterbim/wincreatorStop your AI agent from lying about tests. Hierarchical engineering loops with machine-checkable proof gates.
asimons81/hardproofA persistent, risk-aware engineering protocol for Hermes Agent. Software has to earn done.
dicnunz/agentproofFree local proof reports for AI-built apps and public PRs; $149 async outside proof audit.
DomenicMoran/verified-doneClaude Code skills that stop an agent from reporting work as finished before it has been demonstrated on the running system.
raeseoklee/codexusLocal execution harness for OpenAI Codex with durable ledgers, verification gates, memory, and replay-gated skills.
iFurySt/OpenARDNeutral self-hosted registry and toolkit for Agentic Resource Discovery.
everywan-dev/claude-code-engineeringTurn Claude Code into a team that has to prove it. 44 skills, 8 review agents, a validation router and a knowledge layer where nothing is claimed without a check that could have failed.
bhumik154/claim-checkVerify test-count claims in commit messages against real pytest output, as a pre-commit hook, a Claude Code hook, or a standalone CLI.
Kinneyzhang/llm-output-auditAudit long-form LLM output for factual accuracy, hallucination risk, source quality, and actionable edit suggestions.
evidiq/evidiqThe trust layer for the AI agent economy — verify capability, score risk, prove reputation before every AI transaction.
digsarab112/taskwitnessEvidence-first verification for AI coding agents — prove what changed, what was tested, and what passed.
ohm41321/lucidoneVerification-first discipline for coding agents (Claude Code + Codex CLI): 9-rule doctrine, six skills, adversarial reviewer agent, fail-open enforcement hooks. Done is proven by a command, not by opinion.
TrothByte/low-level-skills-trothbyteProduction-grade low-level engineering skills for AI coding agents. 124 verified, source-backed skills (C, C++, Rust, asm, kernel, embedded, Zig, GPU, RE, build systems) with provenance, examples, and evals.
mohitantil3399/-FactCheck_AgentAn AI-powered Fact-Checking Agent built with Python, Gradio, and Google GenAI.
Reedtrullz/ClankerOSLocal-first agent operating system for durable AI coding: task graphs, deterministic context packs, executable delegation, verification evidence, and approval gates.
markudevelop/outcome-fusion-principiaModel fusion for Claude Code — one model builds, a second model (DeepSeek) judges the results. First-principles missions, a proof ledger, and a release gate that won't let your agent stop early.
BAS-More/truth-shieldStop Claude from lying to you. A Claude Code skill that fact-checks every claim against real sources — code files, live docs, web search — and flags anything it can't confirm.
andyyaro/provalumeVerified, git-aware memory for autonomous software agents—tracking what worked, what failed, what is currently true, and the evidence that proved it.
sylvesterkaczmarek/worldstate-checkDeterministic postcondition verification for AI agents and autonomous systems using independent observed state.
RELATED Other topics · full topics ranking →
#claude-code
1,564#ai-agents
1,160#llm
1,071#claude
943#python
806#ai
741#developer-tools
728#mcp
727#codex
521Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology