# vision-reader

> vision-reader — joshualam21-vision-reader. Use this tool when you need to enable language models to understand images without native vision capabilities. The vision-reader converts image regions into text encodings, such as ASCII art and color stats, and supports features like progressive zoom, OCR, and summary overviews. It takes image files as input and outputs text representations, making it ideal for use cases where visual data needs to be processed and analyzed by LLMs.

Canonical page: https://skillsregistry.net/skills/joshualam21-vision-reader  
JSON: https://api.skillsregistry.net/v1/skills/joshualam21-vision-reader

## Description

MCP server that enables LLMs to understand images without native vision by converting image regions into text encodings (ASCII art, grayscale grids, color stats) and supporting progressive zoom, OCR, and overview summaries. Users can load images, get chunk overviews, crop and encode specific regions, and extract text using normalized coordinates.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-08-29

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-08-29

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/lhbaupvn0n)
- **Repository:** <https://github.com/JoshuaLam21/vision-reader>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "joshualam21-vision-reader"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/joshualam21-vision-reader` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/joshualam21-vision-reader/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
