#hallucination-detection
GitHub repositories that have self-applied the topic "hallucination-detection" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #hallucination-detection (top 7 by stars)
Thoth — agentic systematic literature reviews with per-claim citation verification (cite_check). 8-stage LangGraph agent, authenticated MCP server on the official registry, public eval dashboard, Langfuse/OTel tracing. 676 tests, $0/mo.
bhaskargurram-ai/verifydocTrust layer for document→JSON extraction & AI agents: calibrated per-field confidence + source grounding + accept/review abstention on any OCR/VLM. Ships VerifyDocBench, a novel grounding-conditioned conformal method, and an MCP server.
Vedansh5545/llm-shieldbenchTrustworthy AI evaluation tool for testing chatbot safety, reliability, hallucination behavior, privacy risk, and instruction-following quality.
sunnydubey1111/agent-trajectory-sentinelReal-time detection and repair of LLM agent failures — a one-class behavioural monitor at ~200 µs/step, with 2,823 committed traces.
Hert4/LLM-Certainty-ConsistencyBackend-agnostic black-box hallucination & RAG-faithfulness detection for LLMs — Probabilistic Certainty & Consistency (arXiv:2601.02574). Works on MLX / OpenAI / vLLM via token logprobs; no model internals, no training.
brylop/multi-rag-agentRAG multi-agente con validación QA anti-alucinaciones (OpenAI + Pinecone + LangChain). Modo offline reproducible.
Eqqinox/OllamancerFully-local terminal AI agent for Ollama. 34 tools, MCP, local RAG, deterministic hallucination checks. No cloud, no API keys, nothing leaves your machine.
RELATED Other topics · full topics ranking →
#claude-code
1,555#ai-agents
1,156#llm
1,066#claude
936#python
802#ai
737#developer-tools
723#mcp
719#codex
517Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology