CanopiiCanopiiAll serversEnterprise →

October 2026 edition

The MCP Security Index

We independently scanned 31,782 Model Context Protocol servers. 18% scored D or F, 0% ship committed secrets, and 48% declare no authentication at all.

Frozen · captured · scoring rubric v2.9.0 · methodology

Graded D or F
18%no change

5,721 of 31,782 scanned servers carry a failing or near-failing security grade.

Committed secrets
0%no change

Live credentials found in published source — an attacker can use these as-is.

No declared auth
48%no change

No OAuth 2.1 or protected-resource metadata gating who may invoke the server's tools.

Graded A
28%no change

8,887 servers verifiably pass the controls that apply to them.

Grade distribution

The median MCP server scores 86/100. Half the corpus falls between 61 and 90.

31,782 scored servers

D/F trend
A8,887 · 28%B14,360 · 45%
C2,884 · 9%
D4,555 · 14%
F1,096 · 3%
7,873 unverifiable
Score distribution summary
Mean score76.1
Median score86
25th percentile61
75th percentile90
Lowest score11
Highest score100

Popularity does not mean safety

Adoption and security posture move independently. The most-installed servers are not the safest ones.

By GitHub stars

CohortserversAverage scoreGraded D/F
1,000+ stars297
75
18%
100–999 stars1,063
78
14%
10–99 stars2,544
80
11%
Under 10 stars21,710
82
3%

By monthly downloads

CohortserversAverage scoreGraded D/F
100k+/month36
83
8%
10k–100k/month116
79
11%
1k–10k/month1,491
84
7%
Under 1k/month6,905
86
3%

Authentication and live exposure

We reached 3,239 live MCP endpoints. Of those, 3% declared that authentication was required and then served their tool surface to an anonymous caller.

Live endpoints probed
3,239

Servers we connected to and spoke MCP with directly.

Auth declared but not enforced
3%

Of probed endpoints — documented auth that anonymous callers get straight past.

Tool-poisoning markers
1%

Hidden instructions found in tool descriptions, which hijack the calling agent.

Attack surface and ecosystem

More exposed tools means more ways to be wrong. Ecosystem mix shows where the risk concentrates.

By number of exposed tools

CohortserversAverage scoreGraded D/F
No tools6,581
82
4%
1–5 tools9,844
76
23%
6–20 tools9,304
76
25%
20+ tools6,053
70
14%

By package ecosystem

CohortserversAverage scoreGraded D/F
repo / remote-only17,555
69
28%
npm9,992
86
4%
pypi3,580
81
7%
mcpb482
84
5%
nuget115
89
3%
cargo58
86
5%

What the ecosystem fails most

Every control in the catalog, ranked by failure rate within each domain. Failure rates are a share of servers where the control actually applied and could be evaluated.

Code safety

ControlFailure rateFailedWarnedEvaluated
No command-injection sinksGuardcode.no_command_injection
2%
59923827,364
No dynamic code executionGuardcode.no_code_eval
1%
2588027,364
No unsafe deserializationGuardcode.no_unsafe_deserialize
0%
981727,364
No path traversalGuardcode.no_path_traversal
0%
06,36227,364
No SSRF sinksGuardcode.no_ssrf
0%
08,41227,364

Secrets & credentials

ControlFailure rateFailedWarnedEvaluated
No committed secretsGuardsecrets.no_committed_secrets
0%
0027,364
Credentials sourced from environmentsecrets.from_env
0%
011,58327,364

Dependencies & supply chain

ControlFailure rateFailedWarnedEvaluated
No install/post-install scriptsGuardsupply.no_install_scripts
4%
38109,987
Package name not typosquattingGuardsupply.not_typosquatting
0%
10013,547
No known-vulnerable dependenciessupply.no_known_vulns
0%
03,80317,288
Dependencies pinned (lockfile)supply.deps_pinned
0%
021,17527,364
Published with provenancesupply.provenance
0%
07,7459,987
Established maintainersupply.maintainer_established
0%
012,15413,547

Tool integrity

ControlFailure rateFailedWarnedEvaluated
No over-broad / destructive toolsGuardtool.no_destructive_scope
11%
2,657025,201
Tool descriptions free of injection markersGuardtool.no_injection_markers
2%
403025,201
No risky post-publish tool changes (rug-pull)Guardtool.no_rug_pull
0%
251,8569,162
Strict tool input schemastool.schemas_strict
0%
018,84321,359

Auth & transport

ControlFailure rateFailedWarnedEvaluated
Remote endpoints use TLSGuardtransport.uses_tls
0%
0017,048
Authentication declaredauth.declared
0%
015,18217,048
Execution sandboxingdeploy.sandboxed
0%
022,71727,364
Live endpoint reachabledynamic.reachable
0%
000
Authentication enforceddynamic.auth_enforced
0%
000
No bind-all / exposed debugtransport.no_bind_all
0%
03,39227,364

Maintenance & governance

ControlFailure rateFailedWarnedEvaluated
Repository not archivedGuardmaint.not_archived
1%
226025,614
Actively maintainedmaint.actively_maintained
0%
078025,614
Declares a licensegov.declares_license
0%
07,43825,614
Has a security policygov.security_policy
0%
022,38727,364
Signed releasesgov.signed_releases
0%
000
Adoption & popularityreputation.adoption
0%
016,75526,795

How much we could verify

Absence of evidence is not safety. A partially-scannable server cannot present a high score, so confidence caps it.

CohortserversAverage scoreGraded D/F
High (80%+ verified)23,772
83
5%
60–80% verified3,592
78
3%
Low (under 40% verified)4,418
40
100%

A further 7,873 listed servers declare no public repository or their source cannot be retrieved. They are listed without a score rather than given an invented one.

How to read these numbers

How often does the MCP Security Index update?
The live page recomputes from the database continuously; a permanent edition is frozen at the end of every calendar month at /mcp-security-index/YYYY-MM. Servers themselves are re-ingested and re-scanned every six hours, so a score change shows up on the live page the same day.
What does a D or F grade actually mean?
Grades come from a 0–100 score: A is 90 and up, B 75–89, C 60–74, D 40–59, F below 40. A confirmed security flaw caps the score outright — a server with a verified command-injection sink cannot exceed 20 no matter how good the rest of it is — so D and F overwhelmingly indicate a specific, evidenced problem rather than a general lack of polish.
Are month-over-month deltas comparing the same servers?
No. Distribution deltas cover the whole corpus, which grows as new servers are ingested, so part of any shift is composition rather than servers changing. The 'biggest score drops' section is the cohort-stable view: those are specific servers that were re-scanned and scored lower than before.
Why are some listed servers not scored at all?
If a server declares no public repository, or its source cannot be retrieved because it is private, moved, or removed, there is nothing to verify. Those are listed as unverified with no score rather than given an invented number, and they are excluded from every percentage on this page.
What counts as 'no authentication'?
A server that declares no authentication mechanism — no OAuth 2.1, no protected-resource metadata — so nothing gates who may invoke its tools. Separately, among servers whose live endpoints we probed, some declare that auth is required and then serve their full tool surface to an anonymous caller; that is reported as 'auth declared but not enforced'.
Can I use these numbers in my own work?
Yes. The full dataset behind each edition is downloadable as JSON and CSV from the page itself. Cite the Canopii Trust Index and link to the specific monthly edition you used, so the numbers you quoted remain verifiable.

Take the data

Every number on this page is downloadable. Attribution to the Canopii Trust Index is all we ask.

Looking for a specific server? Browse the full directory — every server page shows its complete control checklist with the evidence behind each result.