AI開発影響研究所 EN
← 全ランキング · 🧠 LLMプロバイダー · 2026-08
LLMプロバイダー

vLLM

vLLM は LLMプロバイダーカテゴリの構成要素として、GitHub上のAI関連Repoでの言及/利用状況を追跡しています。

27
Repoが言及
#13
LLMプロバイダー内ランク
0.6%
占有率
27
取得Repo

REPOS vLLM を使っている / 言及している Repo (TOP 27 / Stars降順)

deepelementlab/jupyter-studio

The AI-native JupyterLab — open-source Cursor for notebooks. Cmd+K inline edit, multi-step agent with cell-level tools (read/edit/run), chat with @cell/@file context, ghost-text completion, one-click traceback fix. BYO model: Anthropic, OpenAI, Gemini, Ollama, vLLM. Local-first, privacy-first, fully open source.

TypeScript 53 AI 50
ding7015869-alt/agent-meeting-studio

🏛️ Multi-Agent Debate & Brainstorm Studio | 本地多Agent辩论&头脑风暴 — run Hermes/Codex/Ollama on your machine

JavaScript 7 AI 100
GrokBuildMJW/ironclad

Reliability for LLM agents through enforcement, not model size — Agent-Contract-Kernel + a fail-closed orchestration engine. Model-agnostic. 🇦🇪 Built in the UAE.

Python 4 AI 70 1 sig
awdemos/toks-bench

Reproducible token-throughput benchmark for OpenAI-compatible LLM servers, tuned for NVIDIA Spark and GB10 inference.

Python 2 AI 50
epsilonagentx/intel_arc_gpu_llm

Docker Compose stack for serving a local, OpenAI-compatible LLM (vLLM on Intel XPU) on an Intel Arc Pro B60 GPU — reproducible config with operator and developer docs.

Shell 2 AI 50
OneCielAI/claude-any

Claude Code provider selector for Anthropic, Ollama, Ollama Cloud, vLLM, NVIDIA hosted, and self-hosted NIM

Python 2 AI 80
dosmoon/aistack
Python 1 AI 50
faizan007jr/local-llm-delegate

Claude Code plugin that delegates simple, well-scoped tasks to a locally running LLM (Ollama, LM Studio, llama.cpp, vLLM) via MCP

JavaScript 1 AI 100
RamazanKara/private-ai-platform-kit

Local-first Kubernetes platform for private LLM and coding-agent workloads.

Python 1 AI 100
wpalish/petrel-rag-v5

On-Premise RAG (продвинутая версия): vLLM/Ollama + bge-m3 + reranker + Qdrant + hybrid (BM25+RRF) + PDR + Basic Auth + Prometheus/Grafana + локальная оценка. Запускается на Ollama без GPU. Тех-задача Petrel AI (Astana Hub).

Python 1 AI 100
manishklach/k3-inference-platform

Production-oriented Kimi K3 inference control plane with checkpoint release gates, MoE capacity planning, OpenAI gateway, NVL72 deployment, benchmarks, and observability.

Python 1 AI 60
gapilongo/pentest-copilot

Self-hosted pentest copilot. Substrate-first (technique catalog + playbooks + RAG + deterministic tools) with structural verifier rules and LLM-as-judge quality eval. Apache 2.0.

Python 1 AI 100
yourself-q/local-browser-agent

Local-first browser agent — attaches to existing Chrome via CDP, runs fully on local LLMs (LM Studio, Ollama, vLLM). Loop detection, multi-action chaining, data-agent-ref grounding.

TypeScript 1 AI 100
Deep-AI-Evo/qwen3.8-27b-q6k-fp8-rtx-pro5000-serving-benchmark

Qwen3.8-27B serving benchmark on RTX PRO 5000: llama.cpp Q6_K vs vLLM FP8/NVFP4 (TTFT/prefill/decode/concurrency)

Python 1 AI 45
msradam/xk6-llm

Load test LLM inference servers with k6. TTFT, ITL, TPOT, goodput, cost, and energy metrics for any OpenAI-compatible server. Ships to Prometheus and Grafana.

Go 0 AI 90
Hert4/LLM-Certainty-Consistency

Backend-agnostic black-box hallucination & RAG-faithfulness detection for LLMs — Probabilistic Certainty & Consistency (arXiv:2601.02574). Works on MLX / OpenAI / vLLM via token logprobs; no model internals, no training.

Python 0 AI 90
eagle-42/askable

Agent-Ops is a technical exploration project focused on instrumentation, evaluation, and GPU serving of LLM agents using OpenTelemetry, Langfuse, RAGAS, and vLLM. It targets generative AI Tech Lead roles in banking and aims to build a robust, long-term LLMOps positioning.

0 AI 45
indiser/CivSim

Turn-based geopolitical simulator where 5 AI civilizations — militarist, mercantile, theocratic, democratic, and authoritarian — reason via LLM (Llama 3.3 70B / Groq) to form alliances, wage wars, conduct espionage, and pursue competing victory conditions. FastAPI game engine · Flask frontend · fantasy SVG world map · injectable event cards.

JavaScript 0 AI 50
Akshitha024/multi-tenant-llm-router

FastAPI multi-LoRA router on top of vLLM: per-tenant auth, rate limits, LRU adapter cache, cost accounting

Python 0 AI 50
cyberlife-coder/llm-local

Thin, zero-dependency CLI to run local LLMs on Apple Silicon (vllm-mlx or mlx_lm) with OpenAI- and Anthropic-compatible endpoints — point Claude Code at a local model.

Python 0 AI 70 2 sig
monthop-gmail/llm-gateway

OpenAI-compatible LLM gateway — LiteLLM + Open WebUI ออก API token เองได้ ต่อ HuggingFace / vLLM / Ollama / cloud providers

Shell 0 AI 65
stpcoder/here-context-recall

Here — 끊긴 업무의 시작점과 다음 행동을 복원하는 데스크톱 앱 · OpenAI-compatible/vLLM

TypeScript 0 AI 60
kevinbtalbert/Claude-Workbench-with-CAI-Inference

Claude Workbench using CAI Inference Service vllm hosted models

Shell 0 AI 75
VAKEELRAKESH/agentos-amd-ai-platform
Python 0 AI 50
luongnv89/dgx-spark-llm-lab

Benchmark local coding LLMs on an OpenAI-compatible endpoint, then keep the serving config that won. Hidden executable tests, reference-validated tasks, mermaid reports.

Python 0 AI 90
Wenri/TaskSolver

Provider-agnostic VLM query flow: one Agent that dispatches to OpenAI / Anthropic / Gemini / vLLM / Claude Code CLI / local HuggingFace model backends, returning parsed answers. Used by 3D-CoT.

Python 0 AI 50 1 sig
Bazaarlinkorg/bazaarlink-byoc-agent

BazaarLink BYOC Agent — connect your own GPU (Ollama, LM Studio, vLLM, llama.cpp) to your own BazaarLink account. MIT.

TypeScript 0 AI 65

※ 「言及」は description / topics / READMEのAI要約の中に "vLLM" の文字列が出現するRepoを示します。

RELATED 同カテゴリの他項目 · LLMプロバイダーの全ランキングを見る →

OpenAI

1,124

Gemini

1,014

Anthropic

686

Ollama

384

Groq

371

DeepSeek

329

Mistral

86

Hugging Face

56

Bedrock

44

EXPLORE 他のカテゴリも見る

⌨️

AIコーディングツール

💻

プログラミング言語

🏷️

GitHub Topic

🧩

AIフレームワーク

🌐

Web/アプリフレームワーク

☁️

クラウド/ホスティング

🔐

認証サービス

🗄️

ベクトルDB

💾

一般データベース

🤖

LLMモデル

🔢

埋め込みモデル

🤝

エージェントフレームワーク

集計対象は AI関連スコア40以上のRepoの最新content snapshot。 算出方法