mcpmark — security grade SAFE, quality 70/100

Security audit verdict: SAFE · quality 70/100

No red flags found in any of the 11 categories — no credential harvesting, no data exfiltration, no curl-pipe-shell installer. Scanned against the SlowMist agent-security taxonomy, refreshed every 8 hours. Full audit →

by eval-sys · MCP Server · ★ 457

Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h

🔒 Is mcpmark safe to install? View the security audit →

About mcpmark

MCPMark: Stress-Testing Comprehensive MCP Use An evaluation suite for agentic models in real MCP tool environments (Notion / GitHub / Filesystem / Postgres / Playwright). MCPMark provides a reproducible, extensible benchmark for researchers and engineers: one-command tasks, isolated sandboxes, auto-resume for failures, unified metrics, and aggregated reports. 🚀 MCPMark Verified is now the default. The standard tasks in this repository are the Verified set — every environment version-pinned and every verifier stabilized. Results from earlier task versions are deprecated and not directly comparable, so please report new numbers as MCPMark Verified. On the Verified set, (xhigh) leads at

agenticbenchmarkeval-sysmcpmcp-serversmcpmarktool-use

Quick Facts

Stars457
Forks42
LanguagePython
CategoryMCP Server
LicenseApache-2.0
Quality Score69.7190947547699/100
Open Issues19
Last Updated2026-06-12
Created2025-07-01
Platformsmcp, python
Est. Tokens~13k

mcpmark alternative? Top 6 similar tools

Looking for a mcpmark alternative? If you're comparing mcpmark with other mcp server tools, these 6 projects are the closest alternatives on Agent Skills Hub — ranked by topic overlap, star count, and community traction.

  • forgemax by postrv · ⭐ 151

    Code Mode inspired local sandboxed MCP Gateway - collapses N servers x M tools into 2 tools (~1,000 tokens)

  • toolhive by stacklok · ⭐ 2.2k

    ToolHive is an enterprise-grade platform for running and managing Model Context Protocol (MCP) servers.

  • mcp-router by mcp-router · ⭐ 2.1k

    MCP Router — development, support and security updates ended 2026-09-18. Historical source and final release f

  • mcpso by chatmcp · ⭐ 2.1k

    directory for Awesome MCP Servers

  • mcphub.nvim by ravitemer · ⭐ 1.8k

    An MCP client for Neovim that seamlessly integrates MCP servers into your editing workflow with an intuitive i

  • tools by strands-agents · ⭐ 1.3k

    A set of tools that gives agents powerful capabilities.

More MCP Server Tools

Explore other popular mcp server tools:

View all MCP Server tools →

Popular Python Agent Tools

Frequently Asked Questions

What is mcpmark?

mcpmark is MCPMark is a comprehensive, stress-testing MCP benchmark designed to evaluate model and agent capabilities in real-world MCP use.. It is categorized as a MCP Server with 457 GitHub stars.

What programming language is mcpmark written in?

mcpmark is primarily written in Python. It covers topics such as agentic, benchmark, eval-sys.

How do I install or use mcpmark?

You can find installation instructions and usage details in the mcpmark GitHub repository at github.com/eval-sys/mcpmark. The project has 457 stars and 42 forks, indicating an active community.

What license does mcpmark use?

mcpmark is released under the Apache-2.0 license, making it free to use and modify according to the license terms.

What are the best alternatives to mcpmark?

The top alternatives to mcpmark on Agent Skills Hub include forgemax, toolhive, mcp-router. Each offers a different approach to the same problem space — compare them side-by-side by stars, quality score, and community activity.

How this security grade is produced

Grades come from a rule-based scan built on the SlowMist agent-security taxonomy, covering 11 red-flag categories including credential harvesting, data exfiltration, and curl | sh installers. It is a first-layer scan, not a manual audit — we say so rather than overstate it.

The scale of the problem is documented independently: Liu et al. (2026), in a study of 31,132 agent skills, report that 26.1% contain security vulnerabilities. Our own full-catalog census is published as a citable open dataset.

Sources & who's responsible:

View on GitHub → Browse MCP Server tools