all the models — AI benchmark observatory
← Benchmarks

MCP-Mark

MCP-Mark evaluates LLMs on their ability to use Model Context Protocol (MCP) tools effectively, testing tool discovery, selection, invocation, and result interpretation across diverse MCP server scenarios.

id mcp-mark · max 1 · 9 models reported

#ModelScore
1Kimi K2.7 Code
Moonshot AI · open
0.81