# Vision MCP Server

> Vision MCP Server — gerardodellan-vision-mcp-server. Use this tool when you need to analyze and understand visual data, such as images and videos, to extract meaningful information. The Vision MCP Server provides multimodal vision tools, including image description, OCR, visual Q&A, and object detection, solving problems like data extraction, content moderation, and visual insights generation. It accepts image or video inputs and outputs relevant data, such as text descriptions or detected objects, making it ideal for applications requiring AI-powered visual analysis.

Canonical page: https://skillsregistry.net/skills/gerardodellan-vision-mcp-server  
JSON: https://api.skillsregistry.net/v1/skills/gerardodellan-vision-mcp-server

## Description

A Model Context Protocol server that provides multimodal vision tools such as image description, OCR, visual Q&A, and object detection, powered by any vision model via OpenRouter.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-09-03

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-09-03

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/udbmcu7tfy)
- **Repository:** <https://github.com/gerardodellan/vision-mcp-server>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "gerardodellan-vision-mcp-server"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/gerardodellan-vision-mcp-server` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/gerardodellan-vision-mcp-server/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
