by Accenture · MCP Server · ★ 497
Last updated: · Indexed by AgentSkillsHub · Auto-synced every 8h
🔒 Is mcp-bench safe to install? View the security audit →
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers Overview MCP-Bench is a comprehensive evaluation framework designed to assess Large Language Models' (LLMs) capabilities in tool-use scenarios through the Model Context Protocol (MCP). This benchmark provides an end-to-end pipeline for evaluating how effectively different LLMs can discover, select, and utilize tools to solve real-world tasks. News [2025-09] MCP-Bench is accepted to NeurIPS 2025 Workshop on Scaling Environments for Agents. Leaderboard | 1
| Stars | 497 |
| Forks | 68 |
| Language | Python |
| Category | MCP Server |
| Quality Score | 68.6570170312/100 |
| Open Issues | 32 |
| Last Updated | 2025-10-07 |
| Created | 2025-08-27 |
| Platforms | mcp, python |
| Est. Tokens | ~1523k |
Explore other popular mcp server tools:
mcp-bench is MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers. It is categorized as a MCP Server with 497 GitHub stars.
mcp-bench is primarily written in Python.
You can find installation instructions and usage details in the mcp-bench GitHub repository at github.com/Accenture/mcp-bench. The project has 497 stars and 68 forks, indicating an active community.