#llama-cpp
GitHub repositories that have self-applied the topic "llama-cpp" — a creator-tagged metadata that surfaces how AI projects describe themselves.
REPOS Repos for #llama-cpp (top 16 by stars)
Lightweight agent that pulls and runs LLM benchmarks from Protorikis Bench
holotherapper/lmmLocal AI model manager
apgamerinfo/boyser-aiCLI coding agent สไตล์ Claude Code — Claude API / Cloud / Ollama / llama.cpp, 42 skills, single-file
vilaca/factory🏭 Coding agent CLI engineered to make any LLM — local or cloud, frontier or 7B — actually useful. 16 providers, multi-tab sessions, automatic key+model failover, plan-mode review, MCP, and recovery loops for models that misbehave.
PavelLizunov/suflyorNative Windows meeting overlay with real-time transcription and local-or-cloud AI answers. Pure Rust + Slint.
awdemos/toks-benchReproducible token-throughput benchmark for OpenAI-compatible LLM servers, tuned for NVIDIA Spark and GB10 inference.
Fandrarista0/opencode-llama-local-agentOne-Click Private Agentic Coding Launch Tool with Local llama.cpp for Developer AI in 2026
ovadmani-sudo/Adaptive-LLM-SamplingBoost LLM reliability with dynamic sampling, retry logic, and a lightweight Go‑based proxy.
StanTheGorilla/GlyphA free, local, open-source alternative to Wispr Flow. Hold a hotkey, talk, and your words are typed into any Windows app — on-device Whisper/Nemotron speech-to-text with optional local LLM cleanup. No cloud, no account.
etherled/local-mcp-searchWindows-first local MCP search for Codex and Claude Code, with local embedding, reranking, and context compression.
velle999/chiron-smacxChiron Rising — an Alpha Centauri (SMACX) mod pack: LLM-written faction diplomacy with leader memory, generated base names, Planetnet dispatches, probe protests, three new factions, and optional Future Society repricing. Built on Thinker.
essentialols/48gb-vram-llm-guideThe 48GB VRAM Local LLM Guide: real benchmarks, model configs, and gotchas from 135+ community reports. Includes runnable benchmark scripts.
eroslifestyle/ai-router-switchSelf-hosted routing proxy for Claude Code and any Anthropic-format client: nine modes across Anthropic, MiniMax, GLM/z.ai, Qwen and local models. One endpoint, per-chat isolation, no IDE restarts.
babatonga/wow-ai-log-analyzerSelf-hosted AI log analyzer for World of Warcraft raid + Mythic+ logs from warcraftlogs.com — actionable improvement reports via Anthropic Claude, OpenAI, or your own llama.cpp GPU.
joshuaswarren/sovereign-inferenceProvider-neutral access & supply layer for open AI: run open models locally (SIN) and route paid, private, verifiable inference across decentralized providers (SIP-AI). DecentralizeAI hackathon entry.
ShAInyXYZ/CerveauLocal-first agentic coding harness — Go + Svelte + llama.cpp, engineered for small local MoE models on consumer hardware. cerveau.sh
RELATED Other topics · full topics ranking →
#claude-code
1,564#ai-agents
1,160#llm
1,071#claude
943#python
806#ai
741#developer-tools
728#mcp
727#codex
521Aggregated by case-insensitive match against topics_json of each repo's latest content snapshot. methodology