# com.clauxel.evalscopebench/evalscopebench-mcp

> Use this tool when you need to benchmark and evaluate AI SDK performance remotely. It provides a paid MCP solution that returns verdicts, receipts, and usage logs, helping to solve problems related to AI model optimization and performance comparison. Ideal for use cases requiring detailed analytics and logging of AI model evaluations, with inputs including AI models and test datasets, and outputs including performance metrics and usage reports.

Canonical page: https://skillsregistry.net/skills/com-clauxel-evalscopebench-evalscopebench-mcp  
JSON: https://api.skillsregistry.net/v1/skills/com-clauxel-evalscopebench-evalscopebench-mcp

## Description

A paid remote MCP for AI SDK benchmark dashboard, built to return verdicts, receipts, usage logs, an

## Trust

- **Trust score (0–1):** 0.30
- **Verification tier:** unverified
- **Last scanned:** 2026-09-03

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** data-analytics
- **Updated:** 2026-09-03

## Source

- **Source listing:** [MCP Registry](https://registry.modelcontextprotocol.io/v0/servers/com.clauxel.evalscopebench%2Fevalscopebench-mcp)
- **Repository:** <https://github.com/clauxel/evalscope-benchmark-mcp-mcp>

## Use it

MCP endpoint published by the skill: `https://evalscopebench.clauxel.com/mcp`

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "com-clauxel-evalscopebench-evalscopebench-mcp"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/com-clauxel-evalscopebench-evalscopebench-mcp` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/com-clauxel-evalscopebench-evalscopebench-mcp/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
