#vision-language-model
GitHub Topic「vision-language-model」がついているAI関連リポジトリの集計。Topicはリポジトリ作者が自己申告するメタタグで、AI関連の文脈で何が「vision-language-model」と分類されているかを可視化します。
REPOS #vision-language-model のRepo (TOP 8 / Stars降順)
Turn documents into AI-ready Markdown with visual understanding
yeahhe365/WebDroid-AgentBrowser-based Android phone agent using WebADB/WebUSB and OpenAI-compatible vision models
CodeChildCZJ/PriorTRPriorTR (ECCV 2026): training-free, prior-corrected visual token reduction for accelerating multimodal LLMs — image & video.
hamadou-08/roboclaw-reportsAI Robotics Demos 2026 - VLM Policies, MCP Skills & HTML Reports
sfyyy/dsh-vision-bridgeOn-demand vision for text-only DeepSeek Harness (DSH) sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model
ai4imaging/MorphAgentAgentic AI for biologically grounded morphological feature design from microscopy — designs, implements, and validates compact cell-profiling features. Includes a desktop Qt UI demo.
pmbstyle/glm-cellphoneLocal HTTP service for running AutoGLM phone-agent tasks against a connected Android device
DiogoRibeiro7/agentic-qa-labAutonomous UI/game-testing agent: vision-language reasoning, browser control, action planning, failure recovery, and evaluation.
RELATED 他のTopicも見る · 全Topicランキング →
#claude-code
1,564#ai-agents
1,160#llm
1,071#claude
943#python
806#ai
741#developer-tools
728#mcp
727#codex
521集計対象: 各Repoの最新contentスナップショットの topics_json に小文字一致でマッチしたもの。 算出方法