iris-eval/mcp-server (iris-eval/mcp-server) is an MCP server listed on the M8ven Trust Index. It scores 74 out of 100, grade C. It declares 9 tools. No publisher has claimed this listing.

C
Emerging
74/100
18 days ago

iris-eval/mcp-server

MCP-native agent evaluation and observability server. Log traces, evaluate output quality with 12 built-in rules (PII detection, prompt injection, cost thresholds), and track agent costs. Real-time dashboard, OTel-compatible spans. Self-hosted, MIT licensed.

Emerging. No concerning findings. Grades remain capped until the project builds reputation through adoption. Grades reflect the full trust pyramid: code, verification depth, and reputation. New projects cap at C until adoption is earned.

How we verified

Code Verified⚡ Live Monitored: not connected

Verified is a snapshot. Live keeps it current, and builds your track record.

⚡ Connect GitHub → continuous verification on every pushwhy connect →

Who stands behind it

iris-eval

Source: Glama

Is this your MCP?

Claim it to get a verified publisher badge, a free copy of our full audit findings, and direct contact for any high-priority issues we find. Or connect your repo for our deepest verification, Live Monitored: read-only, revoke anytime. What we access →

Install from

M8ven verifies MCPs across every public registry — install directly from whichever one you prefer.

// key findings
🚨
Secret credentials may flow to a network call
2 flows detected: DEVTO_API_KEY. We can’t prove the destination matches the brand the credential belongs to.
🔐
You'll be asked for 6 credentials: ANTHROPIC_API_KEY, DEVTO_API_KEY, IRIS_ANTHROPIC_API_KEY, IRIS_API_KEY, IRIS_OPENAI_API_KEY, OPENAI_API_KEY
These are read from process.env at runtime. Make sure you trust where they’ll be sent.
// environment variables
To run this server yourself, you supply these values. They go in your own MCP client configuration and stay on your machine. The secret label means the value is sensitive, not that the server mishandles it.
🔐 secretANTHROPIC_API_KEY
🔐 secretDEVTO_API_KEY
configDRY_RUN
configIRIS_ALLOWED_ORIGINSComma-separated allowed CORS origins
🔐 secretIRIS_ANTHROPIC_API_KEY
🔐 secretIRIS_API_KEYAPI key for HTTP authentication
configIRIS_CITATION_ALLOW_FETCH
configIRIS_CITATION_DOMAINS
configIRIS_DASHBOARDEnable web dashboard (true/false)
configIRIS_DASHBOARD_PORTDashboard port (default 6920)
configIRIS_DB_PATHSQLite database path
configIRIS_HOSTHTTP transport host (default 127.0.0.1)
configIRIS_LLM_JUDGE_MAX_COST_USD_PER_EVAL
configIRIS_LOG_LEVELLog level: debug, info, warn, error
configIRIS_NO_AUTO_LAUNCH
🔐 secretIRIS_OPENAI_API_KEY
configIRIS_OTEL_ENDPOINT
configIRIS_OTEL_HEADERS
configIRIS_OTEL_SERVICE_NAME
configIRIS_OTEL_TIMEOUT_MS
configIRIS_PORTHTTP transport port
configIRIS_TRANSPORTTransport type (stdio or http)
configMAX_NEW
🔐 secretOPENAI_API_KEY
configXDG_CONFIG_HOME
// quality suggestions

Dependencies

21 dependencies, 1 flagged: @playwright/test

Tool inputs are validated

Only 0/9 tool handlers declare input schemas (0%)

Declare an inputSchema with zod/joi/yup on every tool definition.

Tool handlers catch errors

Only 0/9 tool handlers wrap calls in try/catch (0%)

Wrap each tool handler body in try/catch and return a structured error response.

Shell command execution

7 child_process calls — runs shell commands

Prefer library functions over shell-outs. If you must shell out, ensure all inputs are properly escaped.

Secrets not logged

4 secret values sent to console.log

Redact or omit secret values from log output.

Claim the listing to review these findings one by one and send us a correction where you disagree, straight to the team. Claiming also means we tell you when the grade moves, and reach you first if we find anything urgent.

// full audit trail
The findings above are the summary. The full trail, every check we ran, each deduction, the network hosts observed and the dependency advisories, goes to verified publishers, along with an alert whenever a new one lands. Verified publishers can also review each finding and dispute it in one click. Publisher corrections have sharpened several of our checks this month, because the maintainer knows the codebase better than any scanner.
// improvement guidance — verified publishers only
We have 4 concrete improvements we can share with the publisher of this MCP. Each comes with specific guidance to raise the trust score.
// embed badge in your README
[![M8ven Score](https://m8ven.ai/badge/mcp/iris-eval-mcp-server-1owzx9)](https://m8ven.ai/mcp/iris-eval-mcp-server-1owzx9)
Shows your grade and updates automatically. Prefer no grade? Append ?variant=verified to the badge URL.
commit: 22e8195e701fbf5705c6ac096b60adac6e09dfbf
code hash: 702b453d7f366aab88aba1c3e8cfb511c89a9b64dcaa6c088eb9b87976514100
verified: 8/6/2026, 3:15:34 PM
view raw JSON →
Check MCPs from inside your assistant
Tool Check · MCP

Vetting this one by hand? Tool Check is an MCP that scores other MCPs. Add it once and ask Claude, ChatGPT, or any MCP client to grade a server, surface CVEs, check the publisher, and suggest safer alternatives — before you install.

https://m8ven.ai/api/mcp/tool-check
check_toolsearch_toolscompare_toolsrecommend_alternativescheck_publisherreport_concern
How to add it →Free · no account needed · works in any MCP client