# agent-vision

> agent-vision — yanickxia-agent-vision. Use this tool when you need to analyze visual data, such as images and videos, to extract insights and information. The agent-vision tool solves problems like object detection, image classification, and video frame analysis by leveraging OpenAI-compatible vision models. It takes in image or video files as input and outputs analyzed data, making it ideal for use cases where text agents require visual understanding capabilities.

Canonical page: https://skillsregistry.net/skills/yanickxia-agent-vision  
JSON: https://api.skillsregistry.net/v1/skills/yanickxia-agent-vision

## Description

A lightweight MCP server that enables text agents to analyze images and videos using OpenAI-compatible vision models, with tools for image analysis and video frame extraction.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-08-29

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-08-29

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/refqywh0u0)
- **Repository:** <https://github.com/yanickxia/agent-vision>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "yanickxia-agent-vision"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/yanickxia-agent-vision` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/yanickxia-agent-vision/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
