AI開発影響研究所 EN
← Topicランキング · 2026-08
GitHub TOPIC

#inference

GitHub Topic「inference」がついているAI関連リポジトリの集計。Topicはリポジトリ作者が自己申告するメタタグで、AI関連の文脈で何が「inference」と分類されているかを可視化します。

10
タグ付Repo
11
TOP10合計★
3
AIツール痕跡あり
10
TOP表示数

REPOS #inference のRepo (TOP 10 / Stars降順)

k1n0F/vramsuite

Predictive GPU memory framework for AI inference workflows

Python 4 AI 70
Ar9av/mlx-lm-server

Local MLX tuned models through OpenAI compatible LLM, image generation and audio inference on Apple Silicon — Rust + PyO3 + MLX

Rust 3 AI 50 1 sig
modelmeld/modelmeld

OpenAI- and Anthropic-compatible AI gateway. Capability-based routing. Streaming. BYOK passthrough. No key custody.

Python 2 AI 50 1 sig 公開済 ↗
manishklach/k3-inference-platform

Production-oriented Kimi K3 inference control plane with checkpoint release gates, MoE capacity planning, OpenAI gateway, NVL72 deployment, benchmarks, and observability.

Python 1 AI 60
cfregly/gpu-perf-tune

GPU profiling and optimization SKILLS with bundled MCP server

Python 1 AI 45 2 sig
msradam/xk6-llm

Load test LLM inference servers with k6. TTFT, ITL, TPOT, goodput, cost, and energy metrics for any OpenAI-compatible server. Ships to Prometheus and Grafana.

Go 0 AI 90 公開済 ↗
bystray/gonka-mcp-server

MCP server for the GONKA network: run cheap LLM inference through the server (free trial or your own key), get multi-model second opinions, and compare live prices — for any AI agent.

Python 0 AI 70 個人 公開済 ↗
SAGARCHRY0777/inferno

Production-grade distributed ML inference platform — FastAPI gateway, Redis-backed worker pool with dynamic batching, WebSocket result streaming, and a live ops dashboard. Serves YOLO, Whisper, RAG and text models, with an MCP agent server and streaming chat.

TypeScript 0 AI 70
joshuaswarren/sovereign-inference

Provider-neutral access & supply layer for open AI: run open models locally (SIN) and route paid, private, verifiable inference across decentralized providers (SIP-AI). DecentralizeAI hackathon entry.

Python 0 AI 70
iqureshi123/tinyinfer

An LLM inference engine written from scratch in Python and C++/Metal — tokenizer, forward pass, KV cache, and INT4 quantization implemented by hand. No PyTorch, no llama.cpp.

Python 0 AI 70

RELATED 他のTopicも見る · 全Topicランキング →

#claude-code

1,564

#ai-agents

1,160

#llm

1,071

#claude

943

#python

806

#ai

741

#developer-tools

728

#mcp

727

#codex

521

集計対象: 各Repoの最新contentスナップショットの topics_json に小文字一致でマッチしたもの。 算出方法