MCPMark is a comprehensive, stress-testing MCP benchmark designed to evaluate model and agent capabilities in real-world MCP use.
eval-sys/mcpmark is drawing steady momentum at +0.1 stars/day (11th percentile in the tracked cohort), for a momentum score of 37.8/100.
Low breakout odds over the next 14 days, led by star acceleration (59th pctl). Confidence is high given the available history.
Transparent heuristic · logged for model training
Reconstructed from 123 snapshots — the time-series GitHub’s API doesn’t expose.
Live score, refreshed every 6 hours. Links back to this page.
[](https://breakwave.vercel.app/repo/eval-sys/mcpmark)Percentile rank within the tracked cohort. The score self-calibrates — it answers “accelerating vs. everything else,” not raw size.