A
Trusted
98/100
3 days ago

DataJuicer-data-processing-MCP

Data-Juicer is a one-stop system for processing text and multimodal data, suitable for foundation models. The Data-Juicer MCP server provides data processing operators to assist with tasks such as data cleaning, filtering, and deduplication.

Trusted. Deep verification, no outstanding findings, and an established reputation. Grades reflect the full trust pyramid: code, verification depth, and reputation. New projects cap at C until adoption is earned.

How we verified

Code Verified⚡ Live Monitored: not connected

Verified is a snapshot. Live keeps it current, and builds your track record.

⚡ Connect GitHub → continuous verification on every pushwhy connect →

Who stands behind it

modelscope

Source: modelscope

Is this your MCP?

Claim it to get a verified publisher badge, a free copy of our full audit findings, and direct contact for any high-priority issues we find. Or connect your repo for our deepest verification, Live Monitored: read-only, revoke anytime. What we access →

Install from

M8ven verifies MCPs across every public registry — install directly from whichever one you prefer.

// key findings
No credential exfiltration, no sensitive file access, no obfuscation
Static analysis found nothing flowing your secrets to unexpected places.
Open source with a license and README
Anyone can audit the code, the license is declared, and the publisher documents what it does.
🔐
You'll be asked for 1 credential: DATA_JUICER_EMAIL_KEY
These are read from process.env at runtime. Make sure you trust where they’ll be sent.
// environment variables
To run this server yourself, you supply these values. They go in your own MCP client configuration and stay on your machine. The secret label means the value is sensitive, not that the server mishandles it.
configSERVER_TRANSPORT
configMULTI_MODAL
configUSE_LOCAL_OP
configDJ_OPS_LIST_PATH
configCUDA_DEVICE_MAX_CONNECTIONS
configOMP_NUM_THREADS
config_JAVA_OPTIONS
configPYTHONHASHSEED
configLOCAL_RANK
configMASTER_ADDR
configMASTER_PORT
configRANK
configWORLD_SIZE
configDNNLIB_CACHE_DIR
configANALYZER_FONT
configDATA_JUICER_EMAIL_CERT
🔐 secretDATA_JUICER_EMAIL_KEY
// quality suggestions

Tool annotations

No tools have read-only/destructive annotations

Add readOnlyHint or destructiveHint annotations to every tool so hosts can warn users before invoking.

All four hints declared on every tool

6/6 tools missing one or more hints — func (missing: readOnlyHint, destructiveHint, idempotentHint, openWorldHint); get_global_config_schema (missing: readOnlyHint, destructiveHint, idempotentHint, openWorldHint); get_dataset_load_strategies (missing: readOnlyHint, destructiveHint, idempotentHint, openWorldHint), +3 more. OpenAI's directory rejects tools where any of the four hints are missing or non-boolean.

For every tool, set all four hints (readOnlyHint, destructiveHint, idempotentHint, openWorldHint) to explicit true/false values that match the handler’s actual behaviour.

Claim the listing to review these findings one by one and send us a correction where you disagree, straight to the team. Claiming also means we tell you when the grade moves, and reach you first if we find anything urgent.

// full audit trail
The findings above are the summary. The full trail, every check we ran, each deduction, the network hosts observed and the dependency advisories, goes to verified publishers, along with an alert whenever a new one lands. Verified publishers can also review each finding and dispute it in one click. Publisher corrections have sharpened several of our checks this month, because the maintainer knows the codebase better than any scanner.
// improvement guidance — verified publishers only
We have 2 concrete improvements we can share with the publisher of this MCP. Each comes with specific guidance to raise the trust score.
// embed badge in your README
[![M8ven Score](https://m8ven.ai/badge/mcp/modelscope-data-juicer-qlf68k)](https://m8ven.ai/mcp/modelscope-data-juicer-qlf68k)
Shows your grade and updates automatically. Prefer no grade? Append ?variant=verified to the badge URL.
commit: 0a7d12c889970e0b231c346eec0d50b091d2843d
code hash: f6b4acd124be80b1c76184f17539f89b498b2d20bb45f15c7a6a3b8d141e29e9
verified: 8/16/2026, 9:10:20 AM
view raw JSON →
Check MCPs from inside your assistant
Tool Check · MCP

Vetting this one by hand? Tool Check is an MCP that scores other MCPs. Add it once and ask Claude, ChatGPT, or any MCP client to grade a server, surface CVEs, check the publisher, and suggest safer alternatives — before you install.

https://m8ven.ai/api/mcp/tool-check
check_toolsearch_toolscompare_toolsrecommend_alternativescheck_publisherreport_concern
How to add it →Free · no account needed · works in any MCP client