# vision-mcp

> vision-mcp — sulghu-vision-mcp. Use this tool when you need to recognize and extract text from images or describe image content in natural language. It solves problems such as OCR, image description, and text extraction from various image sources, including local files, URLs, and data URLs. The vision-mcp tool takes image inputs and returns text outputs, making it ideal for applications requiring image recognition and text analysis.

Canonical page: https://skillsregistry.net/skills/sulghu-vision-mcp  
JSON: https://api.skillsregistry.net/v1/skills/sulghu-vision-mcp

## Description

An MCP server for image recognition and OCR via OpenAI-compatible vision APIs, supporting local files, URLs, and data URLs. Enables natural language image description and text extraction.

## Trust

- **Trust score (0–1):** 0.68
- **Verification tier:** scanned
- **Last scanned:** 2026-08-29

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-08-29

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/xumqnl435d)
- **Repository:** <https://github.com/sulghu/vision-mcp>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "sulghu-vision-mcp"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/sulghu-vision-mcp` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/sulghu-vision-mcp/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
