Evaluates RAG outputs on faithfulness, answer relevancy, and context precision using an LLM-as-a-Judge backend. Exposes tools for running evaluations, scoring individual samples, and checking thresholds, enabling CI gating and on-demand assessment via MCP.
Claim it to get a verified publisher badge, a free copy of our full audit findings, and direct contact for any high-priority issues we find.
Install from
M8ven verifies MCPs across every public registry — install directly from whichever one you prefer.
process.env. You'll be asked to provide them before it can run.GEMINI_API_KEY— set =... (https://aistudio.google.com/app/apikey)JUDGE_MODELMIN_FAITHFULNESSMIN_ANSWER_RELEVANCYMIN_CONTEXT_PRECISIONUSE_RAGAS— RAGAS/DeepEval (set =true).[](https://m8ven.ai/mcp/saiarja-llm-eval-mcp-gmbg43)