An MCP server that allows Claude Code to offload mechanical tasks such as summarization, classification, and drafting to a local LLM, reducing API costs while keeping Claude in control of complex reasoning and quality review.
Claim it to get a verified publisher badge, a free copy of our full audit findings, and direct contact for any high-priority issues we find.
Install from
M8ven verifies MCPs across every public registry — install directly from whichever one you prefer.
process.env. You'll be asked to provide them before it can run.LOCAL_LLM_BASE_URL— export ="http://localhost:12434/engines/v1"LOCAL_LLM_MAX_TOKENS— 2048 Default max tokensLOCAL_LLM_MODEL— qwen2.5-coder:7b Model to useLOCAL_LLM_TEMPERATURE— 0.7 Default temperature[](https://m8ven.ai/mcp/semenenkod333-mcp-local-llm-b23o5w)