# io.github.hidai25/evalview-mcp

> Use this tool when you need to perform regression testing for AI agents, ensuring their accuracy and reliability over time. It solves problems related to model drift and performance degradation by comparing outputs against golden baselines. The tool integrates with CI/CD pipelines and supports various AI platforms, including LangGraph, CrewAI, OpenAI, and Claude, accepting model inputs and producing test results as output.

Canonical page: https://skillsregistry.net/skills/io-github-hidai25-evalview-mcp  
JSON: https://api.skillsregistry.net/v1/skills/io-github-hidai25-evalview-mcp

## Description

Regression testing for AI agents. Golden baselines, CI/CD, LangGraph, CrewAI, OpenAI, Claude.

## Trust

- **Trust score (0–1):** 0.00
- **Verification tier:** scanned
- **Last scanned:** 2026-09-19

## Facts

- **Version:** 0.3.1
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** devops-ci
- **Updated:** 2026-09-19

## Source

- **Source listing:** [MCP Registry](https://registry.modelcontextprotocol.io/v0/servers/io.github.hidai25%2Fevalview-mcp)
- **Repository:** <https://github.com/hidai25/eval-view>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "io-github-hidai25-evalview-mcp"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/io-github-hidai25-evalview-mcp` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/io-github-hidai25-evalview-mcp/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
