#llm-evals
GitHub repositories that have self-applied the topic "llm-evals" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #llm-evals (top 5 by stars)
Playbook for PMs shipping AI products with PRDs, evals, HITL, launch gates, cost, and observability.
AnonZ7/agentic-rag-eval-harnessProduction-shaped agentic RAG: LangGraph plan->act->verify agent + hybrid retrieval + guardrails + an eval gate in CI. Provider-agnostic; runs offline with no keys.
ch040602/agent-prompt-injection-zooSource-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.
Redsf/rag-internal-knowledge-chatbotSlack-native internal knowledge chatbot with nightly-reindexed RAG (Pinecone) and source-cited answers. Reference build behind a 50% onboarding-time-reduction case study.
mborges-dev/extraction-evalsReproducible benchmark for LLM-based structured extraction from documents. Compare Claude / GPT / Gemini / open-weight on the same task with cost + latency tracking.
RELATED Other topics · full topics ranking →
#claude-code
1,555#ai-agents
1,156#llm
1,066#claude
936#python
802#ai
737#developer-tools
723#mcp
719#codex
517Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology