Computer Vision Tools
CV-MCP-Tools provides a suite of computer vision capabilities for language models through the Model Context Protocol. The repository includes three main components: an image generation server using FLUX.1-Schnell, an OCR server leveraging Qwen-VL and Janus models for text extraction and image understanding, and an object detection tool built on YOLO. Each component is containerized with Docker for easy deployment and exposes APIs for seamless integration. The implementation supports both Claude Desktop and Ollama through configuration files, with MinIO integration for image storage and retrieval. This toolset enables AI assistants to perform complex visual tasks including generating images from text prompts, extracting text from images, and identifying objects in photos without leaving the conversation interface.
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately. Last scanned 2026-09-02.
Scan details: Circle-IR · 2026-09-02 · Appeal
View full trust & usage report →Metadata
- Version
- 1.0.0
- Skill type
- atomic
- Execution layer
- mcp-remote
- Category
- cloud-infra
- Source
- PulseMCP
- Repository
- github.com/omidsrezai/cv-mcp-tools
- Author type
- human
- Last scanned
- 2026-09-02
- Updated
- 2026-09-02
Use via MCP
Resolve Computer Vision Tools from your agent
Streamable HTTP transport at https://api.skillsregistry.net/mcp. No auth for read tools. Discovery: .well-known/mcp.json.
One command in your shell — Claude Code wires it up and verifies the connection. Run /mcp in any session to confirm.
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp --scope user for --scope project to commit it to .mcp.json.