# Vision OCR MCP

> Vision OCR MCP — wbwbwb123-mcp-visoin-ocr. Use this tool when you need to extract text from images or summarize image content, and you have either a file path or a clipboard image as input. Vision OCR MCP solves problems such as converting scanned documents or images of text into editable text, and optionally provides a summary of the image content. It outputs extracted text and optional summary, making it ideal for automating data entry, document processing, and image analysis tasks.

Canonical page: https://skillsregistry.net/skills/wbwbwb123-mcp-visoin-ocr  
JSON: https://api.skillsregistry.net/v1/skills/wbwbwb123-mcp-visoin-ocr

## Description

Local OCR (RapidOCR) plus optional Qwen-VL for image summarization, supporting both file paths and clipboard images.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-09-03

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-09-03

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/bldxnxe2hr)
- **Repository:** <https://github.com/wbwbwb123/mcp-visoin-ocr>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "wbwbwb123-mcp-visoin-ocr"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/wbwbwb123-mcp-visoin-ocr` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/wbwbwb123-mcp-visoin-ocr/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
