AI Dev Impact Lab JA
← Topics ranking · 2026-08
GITHUB TOPIC

#llm-evals

GitHub repositories that have self-applied the topic "llm-evals" — a creator-tagged metadata that surfaces how AI projects describe themselves.

5
tagged repos
2
top 5 stars
0
with tool sigs
5
shown

REPOS Repos for #llm-evals (top 5 by stars)

zaidazmi/AI-PM-PLAYBOOK

Playbook for PMs shipping AI products with PRDs, evals, HITL, launch gates, cost, and observability.

TypeScript 2 AI 100 Solo live ↗
AnonZ7/agentic-rag-eval-harness

Production-shaped agentic RAG: LangGraph plan->act->verify agent + hybrid retrieval + guardrails + an eval gate in CI. Provider-agnostic; runs offline with no keys.

Python 0 AI 100
ch040602/agent-prompt-injection-zoo

Source-backed archive of agent prompt-injection incidents, research records, trust-boundary patterns, schemas, and sanitized defensive summaries.

HTML 0 AI 100 Solo live ↗
Redsf/rag-internal-knowledge-chatbot

Slack-native internal knowledge chatbot with nightly-reindexed RAG (Pinecone) and source-cited answers. Reference build behind a 50% onboarding-time-reduction case study.

Python 0 AI 100
mborges-dev/extraction-evals

Reproducible benchmark for LLM-based structured extraction from documents. Compare Claude / GPT / Gemini / open-weight on the same task with cost + latency tracking.

Python 0 AI 70

RELATED Other topics · full topics ranking →

#claude-code

1,555

#ai-agents

1,156

#llm

1,066

#claude

936

#python

802

#ai

737

#developer-tools

723

#mcp

719

#codex

517

Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology