Read-only ops/health MCP server for a local vLLM instance: health tiers, GPU/VRAM, service status.
Claim it to get a verified publisher badge, a free copy of our full audit findings, and direct contact for any high-priority issues we find.
Install from
M8ven verifies MCPs across every public registry — install directly from whichever one you prefer.
process.env. You'll be asked to provide them before it can run.VLLM_OPS_MCP_BASE_URL— OpenAI-compatible base URL http://127.0.0.1:8000/v1 (explicit 127.0.0.1, not localhost -- see qwen_cli.py's IPv6-hang footgun note)VLLM_OPS_MCP_MODEL— served-model-name, used in completion payloads qwen3-14bVLLM_OPS_MCP_WSL_DISTRO— WSL distro hosting vLLM Ubuntu-22.04VLLM_OPS_MCP_SERVICE_UNIT— systemd unit name vllmVLLM_OPS_MCP_EXEC_SCRIPT_PATH— WSL-side path to the serve exec script (fallback source) ~/vllm-systemd-exec.shVLLM_OPS_MCP_NVIDIA_SMI_PATH— nvidia-smi binary override nvidia-smi (Windows host) / nvidia-smi.exe (native WSL2)VLLM_OPS_MCP_SERVE_CONFIG_PATH— local file mirror of the exec script; if set, get_serve_config reads it directly instead of shelling into WSL as a fallback source unsetVLLM_OPS_MCP_RUNS_IN_WSL— 1 if this server process itself runs insideVLLM_OPS_MCP_DEEP_RATE_LIMIT_PER_MIN— ) specifically because this vLLMVLLM_OPS_MCP_SCRATCH_DIR[](https://m8ven.ai/mcp/jaimenbell-vllm-ops-mcp-adnfr7)